Pith. sign in

Paper Citation Record · LEDGER

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

As of 13 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2411.14704.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14704 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:57.863886Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T00:02:12.727566Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T00:29:47.839616Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4c3c0f33-d76b-44b2-828f-248f4bb5975a · outbound

This paper cites Remote sensing big data computing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing big data computing: Challenges and opportunities,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.612459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.642747Z digest=sha256:caf3085f1d89c7cdb969afeca768da8bf27fe6d39a7871ec7ebcb54fe75e2794

Observation b520db60-7253-48a9-ab56-357ceba5605a · outbound

This paper cites Big data for remote sensing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Big data for remote sensing: Challenges and opportunities,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.599090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.648140Z digest=sha256:cad551714d196092cc31277565ba4efd5908995684149967ea75899c820a64ba

Observation 186ad7d1-a193-4143-a9b6-b76fada333c3 · outbound

This paper cites Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.583419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.653137Z digest=sha256:047e71517bd65f7067767491c2f88e02f45e31d4dd76a9caa9bfd9cf3d0bfa7d

Observation 43844814-7191-468d-a558-7acd3ab66674 · outbound

This paper cites Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.569310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.660856Z digest=sha256:f3ae3ec044a2fa04edebea3565d6a663bfd5e8fc546c6a4ebfd4550b46388e39

Observation 6fa11a44-fe85-41c9-a49a-5b0c45c4ce5f · outbound

This paper cites Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.554913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.665468Z digest=sha256:63e0aeddafad3b050782be059eb26e50cfbfc3b3d480871e1e748cb20eae78b1

Observation 960bfcb1-ca03-4497-8df8-27754115cfda · outbound

This paper cites Nwpu- captions dataset and mlca-net for remote sensing image captioning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Nwpu- captions dataset and mlca-net for remote sensing image captioning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.669876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.669876Z digest=sha256:99161f8d025d28bb04b7dceafb634ac40abcf6ebfdbe4ef44736b3dd1a485da6

Observation 654ef5dd-7f5b-4121-bd63-54b3ae71712e · outbound

This paper cites Textrs: Deep bidirectional triplet network for matching text to remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Textrs: Deep bidirectional triplet network for matching text to remote sensing images,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.521644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.674957Z digest=sha256:8a012f18700960056aee8bdad7d0c0601e4eb961e6ab64cff5a34e5627ced47c

Observation 313642ac-f437-493f-823a-de51ef14e235 · outbound

This paper cites A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.679755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.679755Z digest=sha256:bdd4592fffba16b763d572301033544aa5a270d24f823b11f13f59b1da8dd3df

Observation 0a042db3-5854-4d0e-9c98-8f731f04820f · outbound

This paper cites Fusion-based correlation learning model for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fusion-based correlation learning model for cross-modal remote sensing image retrieval,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.499246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.683688Z digest=sha256:5e9a1c83716ebc8f8d6528804712bb770650083ab497f4f6abd9c3358735ab68

Observation b1447892-a6cf-4ffb-84d5-2d9dfe23f503 · outbound

This paper cites Cross spectral image reconstruction using a deep guided neural network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross spectral image reconstruction using a deep guided neural network,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.486375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.687271Z digest=sha256:145a2ca5843c41c493a005d1bb53c7b417aa9e6d081753eed99706f7768bf91f

Observation f87b27fd-5a3f-4f20-b745-98abb2d47e4f · outbound

This paper cites Image super-resolution using t-tetromino pixels,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image super-resolution using t-tetromino pixels,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.472800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.691404Z digest=sha256:9d0ededc82e151e97c7dcf44f3a0d3f9bf9fa338d22eb1de8068958cb9685530

Observation a2bafe0a-f97a-414e-9580-6e01c0556ead · outbound

This paper cites Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.695460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.695460Z digest=sha256:76eceaecc20ed32f61b3495c9e376574fa0cd0df2ee9659db49f07162a8ccd21

Observation deb1efde-12a0-4211-994d-3d1cc98e2126 · outbound

This paper cites Remote sensing cross-modal text-image retrieval based on global and local information,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing cross-modal text-image retrieval based on global and local information,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.699844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.699844Z digest=sha256:3be4263ea9fa45296ae44e1a5d80311500c7532ef75ff20a07df2bfa2a7f246f

Observation 14cdfa4f-a120-4b0e-ae72-c3c707dcabf0 · outbound

This paper cites A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.440911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.704005Z digest=sha256:8750492776abf3f6e4901a55ee4cb276adaa2e67329c0e78bc98063206e63c02

Observation c8db1e29-8418-4854-ae4c-169435be8548 · outbound

This paper cites Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.427872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.707863Z digest=sha256:0a3db26853dbf705a739ebfed84d921828c37d77b89b8f1246b627040cb9f68a

Observation 70b6b766-acec-470d-b78a-489c3a11e6af · outbound

This paper cites Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.413683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.711725Z digest=sha256:eec825681f6fbab0ef46a395a333f5355dcea552ffc76633613a145ed37567ca

Observation d19e02e5-e545-43e2-b5ad-b9799914777c · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.715976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.715976Z digest=sha256:effda024b4ad0d2cfdf169280417bc1938cb273e834649b7a9476ff27367bf4f

Observation 4f20a153-5c0f-4009-a05c-947e74b345ca · outbound

This paper cites Long short-term memory recurrent neural network architectures for large scale acoustic modeling,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Long short-term memory recurrent neural network architectures for large scale acoustic modeling,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.399313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.720276Z digest=sha256:9fce721ce4a8fcfb42e3a69617afabdf1eb17f52357c1de51e1b774efbc4b93b

Observation beb708e3-24d7-4cce-ae2e-c8810b2f1c7d · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.724413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.724413Z digest=sha256:fc5f1f09b99e5908661eeda54c103e0c78735a93c9beb4b0c736da406d55c7a4

Observation 6dce598e-f33e-42cb-8ffc-3cf4bb0e6501 · outbound

This paper cites Attention is all you need,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Attention is all you need,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.728685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.728685Z digest=sha256:227e4a93c5b651158af3481c29cbb457d1f7cc563652dd6a7c8734e0cbaa7497

Observation dce4e68e-4451-4a12-980a-ede130182e20 · outbound

This paper cites Multiscale salient alignment learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multiscale salient alignment learning for remote sensing image-text retrieval,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.378152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.732957Z digest=sha256:ee172aae5e4f1e8e141509288e3d15f524138f8d943762b7f998b91912265959

Observation 71c63976-4336-441b-8ff7-2f40c3b50181 · outbound

This paper cites Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.364329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.736707Z digest=sha256:34445c7fe8f796223cd49a879c05721b18331047960877ecf0ae8b14f2a483dc

Observation 9465be24-0dca-4812-a1c8-70fdddfadfc5 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Align before fuse: Vision and language representation learning with momentum distillation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.349930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.740423Z digest=sha256:b8248e84430a7d3cc758b2de86973d8387ebed7c268341b09e5766ede8fda7af

Observation e6a07d92-cebb-495a-8826-6c0c53a5ed8c · outbound

This paper cites Deep saliency smoothing hashing for drone image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep saliency smoothing hashing for drone image retrieval,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.332354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.744193Z digest=sha256:02fe0f3f1b4c6c1bda960792b9924dfea12d1803895ed3054fc9ab3789b16fee

Observation 76ed2888-dad4-4d5a-9698-91351679a81c · outbound

This paper cites Multitask learning for sar ship detection with gaussian-mask joint segmentation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multitask learning for sar ship detection with gaussian-mask joint segmentation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.314377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.748050Z digest=sha256:28dad47b8bb03b74f61afc71d58b09043afb572ea60a17b55a6380d1f6c9d752

Observation 930a1883-858a-468c-ae0d-e8757f9c2a18 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.302154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.751703Z digest=sha256:1f497eb1a053d65a524ae89c4a28e3460803c2d83f0e3514af70a5e2c72c7a75

Observation 2380a155-4554-40be-b0f7-8cda192742e4 · outbound

This paper cites Global context vision transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Global context vision transformers,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.289581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.755719Z digest=sha256:ac304e176a9192767be6fb463cc81c734be3d5813ed0e3178e2873551cc1bac8

Observation 37aa4e64-5910-4588-9fbb-cf1ea37451dc · outbound

This paper cites Matching images and text with multi-modal tensor fusion and re- ranking,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Matching images and text with multi-modal tensor fusion and re- ranking,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.276858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.759505Z digest=sha256:a88eb6e04a7ace3fddf29e8b9bdb554d6b1f52fc6ab83705ba1f21ca4a6f8942

Observation 6d79f654-7652-4f77-a091-b4447c4ddbfb · outbound

This paper cites Vse++: Improving visual-semantic embeddings with hard negatives,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vse++: Improving visual-semantic embeddings with hard negatives,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.264258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.763722Z digest=sha256:3e03dd126a84b1d0fad5eb88a70c21a02466525aca329833777245914686e7c5

Observation 84ac96ff-b178-43ae-82b3-893433494fb9 · outbound

This paper cites Exploring models and data for remote sensing image caption generation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring models and data for remote sensing image caption generation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.767188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.767188Z digest=sha256:a656f95feef5907a5b1e90d9daacae30d725a0cbb53e0823a1fcab03b34be020

Observation fafb1176-24c3-427d-b2a1-e7c343df78e3 · outbound

This paper cites Deep semantic understanding of high resolution remote sensing image,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep semantic understanding of high resolution remote sensing image,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.242536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.770808Z digest=sha256:d06467455123258e2650594b3836672f30f82bd1b6b2301a604646c60f31fc0e

Observation 44d0bb0b-68d2-4af5-bf61-64699935b562 · outbound

This paper cites End-to-end convolutional semantic embeddings,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval End-to-end convolutional semantic embeddings,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.225883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.774534Z digest=sha256:ea9465e5fea8ff11c5b4884e417af373ad36870804ea3c670530f2dc7afb19f6

Observation 519cb981-5d09-44c6-896e-a8e9aecc3cd6 · outbound

This paper cites Cross-modal semantic correlation learning by bi-cnn network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross-modal semantic correlation learning by bi-cnn network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.210444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.778748Z digest=sha256:9bd5216091712294b9e392d66b22d9cfd7b93882268c3210ab09fe91210abec9

Observation f110aff3-2c77-467e-856d-caf394a2b75a · outbound

This paper cites Dual-path convolutional image-text embeddings with instance loss,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Dual-path convolutional image-text embeddings with instance loss,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.195285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.782641Z digest=sha256:f68f8f5fe131f9647b4271eae5fd4ae858744dbbc752cc3f6e54210d73f763c8

Observation e1216bf4-dd3e-4a46-a799-90b70d2aa7f3 · outbound

This paper cites Deep supervised cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep supervised cross-modal retrieval,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.786515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.786515Z digest=sha256:15e4daa2a8519125179853059eb5fb6f6037a7c93a3eeae5db1bd895bfbc387f

Observation 03dc5032-001b-4cd2-9d2d-27ef02950545 · outbound

This paper cites Learning semantic concepts and order for image and sentence matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning semantic concepts and order for image and sentence matching,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.170650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.790384Z digest=sha256:f2667edf4a66c79f7e88daaceab7de2c0d36113bf2647c11b8e4a3c072859b36

Observation ee5e49a4-9c8a-4ed9-be2d-5070c4577c11 · outbound

This paper cites Stacked cross attention for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Stacked cross attention for image-text matching,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.156138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.794558Z digest=sha256:9834c3d70b5e2d1bfe95f8c3c0a57f1ee9b7b341c6bedb5438d09960f56ad26f

Observation 8fd228ad-e85b-43c2-a565-2c2e4aa3d91a · outbound

This paper cites Cross- modal attention with semantic consistence for image–text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross- modal attention with semantic consistence for image–text matching,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.142074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.798448Z digest=sha256:205882cc9960cee3bbbcc3e37501446836072fc2d83c88da9c31df77a2157ca7

Observation 38ae2995-86c3-4b14-93b2-da73ff3a6f7f · outbound

This paper cites Visual semantic reasoning for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Visual semantic reasoning for image-text matching,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.129587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.802192Z digest=sha256:aee2c99bdb9c786adae12e860ab122f98f90aae33ff95d284ef55e350030c07f

Observation 91c20eb1-1199-455f-a128-3099ece39e98 · outbound

This paper cites Image-text embedding learning via visual and textual semantic reasoning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image-text embedding learning via visual and textual semantic reasoning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.115153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.805867Z digest=sha256:8a64a8150956850e91b88ba6cb8acacb5616f359ca2018a5ee2546eeef44ce51

Observation 1414cfaf-f4aa-449e-a1a5-f582b8bb01c7 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.809534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.809534Z digest=sha256:02e5a285203e8f6ce48f882d77a1b1b1d4ca199ded39a49da757b1a56acc21fa

Observation ecf46abb-c98e-4236-af91-6bb7d5159778 · outbound

This paper cites Lxmert: Learning cross-modality encoder representations from transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Lxmert: Learning cross-modality encoder representations from transformers,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.093662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.813277Z digest=sha256:8475a359cada42afc08d7b7689becf2559b19b67f5e9952a2a9c82c16a717c5f

Observation c1a80255-d52c-4e74-87fb-1c654dfcc07f · outbound

This paper cites Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.080079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.816997Z digest=sha256:3f7253063f533f89d41c9c78f387b203d041a20e16a7674745ea18ab4917e86d

Observation 84f014d7-3af0-4263-b883-94ef99a55daa · outbound

This paper cites Learning the best pooling strategy for visual semantic embedding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning the best pooling strategy for visual semantic embedding,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.066838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.820850Z digest=sha256:e31c73f5984147ae756db14561ea23dbf843d49dd07cc2f5f05aee11b87298a5

Observation 0e90948c-19b9-4e50-8e10-f5988d1ed44f · outbound

This paper cites Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.824935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.824935Z digest=sha256:1f17638afaed7953cfab4912910f1add303da1ece98d4a25b34216b4f1f22d5e

Observation bf64ed5d-685c-43e1-9734-415474ed2ece · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning transferable visual models from natural language supervision,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.829050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.829050Z digest=sha256:5f4f3064e4cb8b7a0d3a8392b42d458d8b1ea7b0abeb7170e8b6a7e91632db2a

Observation e2c38548-073c-429f-a7d2-c76397f0bc7a · outbound

This paper cites Vista: Vision and scene text aggregation for cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vista: Vision and scene text aggregation for cross-modal retrieval,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.043211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.832696Z digest=sha256:ad252fcefbd1f34254caa193eec690f4ad9748c4ee892a0d194e5489b43d93da

Observation d6f8ddff-9a5c-43b6-82b6-5bc90cca8ffb · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.027761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.836567Z digest=sha256:318cd5159cb718313e04e930e23a89e7ddac189eeeda80c3d8891610cf9b85a6

Observation 597fa9db-e019-4720-8555-7da321fee29a · outbound

This paper cites Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.013236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.840435Z digest=sha256:addf3cbae85f4a00b043c5810b5001fabe4f5ed03d3bfbd0fb7ca88247b5de93

Observation 45903d78-eded-4900-a047-c4855e2cfac6 · outbound

This paper cites Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.999616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.844283Z digest=sha256:a6b5d309dd69b5e51912629fa586f1106784388750a797d4676b816c0b0bd716

Observation 9b3afb81-747a-406f-858f-61a9b7dd3591 · outbound

This paper cites Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.983362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.848175Z digest=sha256:60b5944d16f89b4fa2501023d760afb05dc3fd57cd910d515607fc0765286a36

Observation 87a56034-0031-4957-9722-3c6a38a28750 · outbound

This paper cites Parameter-efficient transfer learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Parameter-efficient transfer learning for remote sensing image-text retrieval,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.852135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.852135Z digest=sha256:48e24a4457678375cf6e7467c4fb9fa6a8dc0060c602816677beb5c72cd83104

Observation 8e1b2a61-fe9a-4da8-bcee-5a7041d2c5ce · outbound

This paper cites Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.856091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.856091Z digest=sha256:0399bd7c79693d99d9f07242cf86c073695c63e1770c624f543092c3b4a55f76

Observation 1ac19980-e07b-4711-8c32-64bc5507a56c · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.951457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.859812Z digest=sha256:116c99fbbee91d4fd96e71fe09577abbc61e2d17da9eb90dfa24cfb3c8a1c896

Observation 852a2f44-d06f-416e-843f-3fb4879354b6 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Momentum contrast for unsupervised visual representation learning,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.937860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T15:05:57.863886Z digest=sha256:64d9e78054f22cfa8d92ff24a1fb776074eecbef1840314984a3f966d0930156

Pith citing papers

Observation be331e05-aba8-4882-8bdd-c298a7fd228c · inbound

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing cites this paper.

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.841047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T00:02:12.727566Z digest=sha256:c55f6353f728fe84efd996c2518928ce9a1ad1dc289618992d06ba7501a17bce