Pith. sign in

Paper Citation Record · LEDGER

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

As of 18 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2411.14704.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14704 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:57.863886Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T00:02:12.727566Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T00:29:47.839616Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4c3c0f33-d76b-44b2-828f-248f4bb5975a · outbound

This paper cites Remote sensing big data computing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing big data computing: Challenges and opportunities,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.612459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.642747Z digest=sha256:e9b912457d3d7ee7734198be3a8902f68bc2bcd1d94416fc22fa24027437b925

Observation b520db60-7253-48a9-ab56-357ceba5605a · outbound

This paper cites Big data for remote sensing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Big data for remote sensing: Challenges and opportunities,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.599090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.648140Z digest=sha256:eb779fc69fc97477b767b001be263e95d284566d3dd480873a2c0b6f4efdf3bb

Observation 186ad7d1-a193-4143-a9b6-b76fada333c3 · outbound

This paper cites Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.583419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.653137Z digest=sha256:8ccb9d86de553e9760b5896a6b494c79f39f8b59690d601e5558ccb15efa77d2

Observation 43844814-7191-468d-a558-7acd3ab66674 · outbound

This paper cites Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.569310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.660856Z digest=sha256:b395223237dc336737fe2c144717b9baaf946fa23c8629cd918b8978a45785be

Observation 6fa11a44-fe85-41c9-a49a-5b0c45c4ce5f · outbound

This paper cites Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.554913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.665468Z digest=sha256:40d3422046a3bde4ea0108244490b47d4889ec160baedf7e0a4fe0f695211d1e

Observation 960bfcb1-ca03-4497-8df8-27754115cfda · outbound

This paper cites Nwpu- captions dataset and mlca-net for remote sensing image captioning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Nwpu- captions dataset and mlca-net for remote sensing image captioning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.669876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.669876Z digest=sha256:067d394abb4c3ef1d2e12077ee4eac563b9fcda15544f97e306127b39cc34c42

Observation 654ef5dd-7f5b-4121-bd63-54b3ae71712e · outbound

This paper cites Textrs: Deep bidirectional triplet network for matching text to remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Textrs: Deep bidirectional triplet network for matching text to remote sensing images,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.521644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.674957Z digest=sha256:906c21432dad377fb079849060e77b20f2296691ae00a21036ca951e8a803e93

Observation 313642ac-f437-493f-823a-de51ef14e235 · outbound

This paper cites A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.679755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.679755Z digest=sha256:a20cd672cf5af24c400e3a1d454ca176ea03a30eb25ff161823c546d04ddce65

Observation 0a042db3-5854-4d0e-9c98-8f731f04820f · outbound

This paper cites Fusion-based correlation learning model for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fusion-based correlation learning model for cross-modal remote sensing image retrieval,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.499246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.683688Z digest=sha256:d09fa3ffa3d8a531addc5517f2ef015acda052e61a0bc26d66ba3fab4a72fdd5

Observation b1447892-a6cf-4ffb-84d5-2d9dfe23f503 · outbound

This paper cites Cross spectral image reconstruction using a deep guided neural network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross spectral image reconstruction using a deep guided neural network,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.486375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.687271Z digest=sha256:2c6ada8406939f26a461cfaa6de83bb3b4a5754e9bb28b7370f65aa426f86779

Observation f87b27fd-5a3f-4f20-b745-98abb2d47e4f · outbound

This paper cites Image super-resolution using t-tetromino pixels,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image super-resolution using t-tetromino pixels,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.472800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.691404Z digest=sha256:3749175d8e02a0af08d5c148b87f939b75631a984036dde6768d1b1a742f1193

Observation a2bafe0a-f97a-414e-9580-6e01c0556ead · outbound

This paper cites Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.695460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.695460Z digest=sha256:39ad7ae7772ec65be220eb50b6411bf7738e828f24d638e975125fe7898045fb

Observation deb1efde-12a0-4211-994d-3d1cc98e2126 · outbound

This paper cites Remote sensing cross-modal text-image retrieval based on global and local information,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing cross-modal text-image retrieval based on global and local information,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.699844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.699844Z digest=sha256:f2b279fdc4c2c6ed71e68d04bc13a769d2d55d1355e7d9bde3ccca3c6d8f3559

Observation 14cdfa4f-a120-4b0e-ae72-c3c707dcabf0 · outbound

This paper cites A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.440911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.704005Z digest=sha256:710744610f92cf79a0cf1f75bfbb4c3cb3c9e39a779c4c7c31c36ed9f674e988

Observation c8db1e29-8418-4854-ae4c-169435be8548 · outbound

This paper cites Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.427872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.707863Z digest=sha256:daea7230b460504e4c8b5f022aa90a10cf32f81e687c0d544120f0b1f1756d34

Observation 70b6b766-acec-470d-b78a-489c3a11e6af · outbound

This paper cites Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.413683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.711725Z digest=sha256:fc155608570b76afd07401609399db9f3e16dbdf40ea61a0c7dd950830b49126

Observation d19e02e5-e545-43e2-b5ad-b9799914777c · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.715976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.715976Z digest=sha256:89fc37addd472f8757cda63f5858bbec50182848c1185fb890fd8a705872e0f4

Observation 4f20a153-5c0f-4009-a05c-947e74b345ca · outbound

This paper cites Long short-term memory recurrent neural network architectures for large scale acoustic modeling,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Long short-term memory recurrent neural network architectures for large scale acoustic modeling,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.399313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.720276Z digest=sha256:17e5896960e9b35eeeb6e4e334996e29c7778c6f7e85f36fb5555e44c41d49ee

Observation beb708e3-24d7-4cce-ae2e-c8810b2f1c7d · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.724413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.724413Z digest=sha256:e33e1907e161b1347508124cd3f0662eac61047e2b484cc9dd9374c772fe04a6

Observation 6dce598e-f33e-42cb-8ffc-3cf4bb0e6501 · outbound

This paper cites Attention is all you need,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Attention is all you need,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.728685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.728685Z digest=sha256:4fb62a4f02498e4af601f546053485fd4154ff3cac472d42d97c1785816e5f8f

Observation dce4e68e-4451-4a12-980a-ede130182e20 · outbound

This paper cites Multiscale salient alignment learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multiscale salient alignment learning for remote sensing image-text retrieval,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.378152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.732957Z digest=sha256:80da6f99d3b18d5175d0771dcf5ccf0cf79b31e59b49d40187d0a2ac1824457f

Observation 71c63976-4336-441b-8ff7-2f40c3b50181 · outbound

This paper cites Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.364329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.736707Z digest=sha256:102fab525ae25260cfa990a8bd6b2ee76b784fb4817c8ed44f76c2be2fa5e1c8

Observation 9465be24-0dca-4812-a1c8-70fdddfadfc5 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Align before fuse: Vision and language representation learning with momentum distillation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.349930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.740423Z digest=sha256:26b0717afe1ab97d4ffdc81abce46ee8f74706f338abe93d8ffabd36abf2c4aa

Observation e6a07d92-cebb-495a-8826-6c0c53a5ed8c · outbound

This paper cites Deep saliency smoothing hashing for drone image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep saliency smoothing hashing for drone image retrieval,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.332354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.744193Z digest=sha256:760460109d2ba1d44d31569b5bb8484a5f1aced33605867ce4896d69134ee2c7

Observation 76ed2888-dad4-4d5a-9698-91351679a81c · outbound

This paper cites Multitask learning for sar ship detection with gaussian-mask joint segmentation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multitask learning for sar ship detection with gaussian-mask joint segmentation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.314377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.748050Z digest=sha256:9c868cd085383673f922045ed59a4551cb2fab04587ee18acc1977e8318f6a7d

Observation 930a1883-858a-468c-ae0d-e8757f9c2a18 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.302154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.751703Z digest=sha256:7a261703fb69bb5c9c8f2702b48065e32cc94234f1c70f2f03c3a7daf8949d26

Observation 2380a155-4554-40be-b0f7-8cda192742e4 · outbound

This paper cites Global context vision transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Global context vision transformers,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.289581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.755719Z digest=sha256:e11385f03d04ce17e95b4798c31b2ad123b5f8ba8d7c635e015ce53215019c09

Observation 37aa4e64-5910-4588-9fbb-cf1ea37451dc · outbound

This paper cites Matching images and text with multi-modal tensor fusion and re- ranking,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Matching images and text with multi-modal tensor fusion and re- ranking,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.276858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.759505Z digest=sha256:650813248fc761c434f0388ffca43647cc00efe7b56ca52ac2c57e7add487ed1

Observation 6d79f654-7652-4f77-a091-b4447c4ddbfb · outbound

This paper cites Vse++: Improving visual-semantic embeddings with hard negatives,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vse++: Improving visual-semantic embeddings with hard negatives,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.264258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.763722Z digest=sha256:a0e82076f3a9d7403d1d0c476b55d7dad417acba30531900e98d9de09b40e125

Observation 84ac96ff-b178-43ae-82b3-893433494fb9 · outbound

This paper cites Exploring models and data for remote sensing image caption generation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring models and data for remote sensing image caption generation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.767188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.767188Z digest=sha256:8a98d03482b86afc8484141046a2a4545d58320037da6cfd231550253f6e69a8

Observation fafb1176-24c3-427d-b2a1-e7c343df78e3 · outbound

This paper cites Deep semantic understanding of high resolution remote sensing image,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep semantic understanding of high resolution remote sensing image,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.242536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.770808Z digest=sha256:f02a1c299f378b51abf2e29d50af14ad2dfeb2adb2683d5d77390037ec4f069f

Observation 44d0bb0b-68d2-4af5-bf61-64699935b562 · outbound

This paper cites End-to-end convolutional semantic embeddings,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval End-to-end convolutional semantic embeddings,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.225883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.774534Z digest=sha256:9c15a0f0b74223e9b98babfb4638445c4d9975377a3293b4d998e3206f349466

Observation 519cb981-5d09-44c6-896e-a8e9aecc3cd6 · outbound

This paper cites Cross-modal semantic correlation learning by bi-cnn network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross-modal semantic correlation learning by bi-cnn network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.210444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.778748Z digest=sha256:1487eacb083a72293d209254b040ab7b653f561cf0ca90f495c2a6929c0d7774

Observation f110aff3-2c77-467e-856d-caf394a2b75a · outbound

This paper cites Dual-path convolutional image-text embeddings with instance loss,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Dual-path convolutional image-text embeddings with instance loss,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.195285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.782641Z digest=sha256:da1ed1a3866a389ef4e6d872248d2de578c6893a2e737988f3e3bd31676013d3

Observation e1216bf4-dd3e-4a46-a799-90b70d2aa7f3 · outbound

This paper cites Deep supervised cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep supervised cross-modal retrieval,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.786515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.786515Z digest=sha256:8b307e3f1220fcf11548992e02595b555365e9d9aff0952d13e6c16d3eb6bd7e

Observation 03dc5032-001b-4cd2-9d2d-27ef02950545 · outbound

This paper cites Learning semantic concepts and order for image and sentence matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning semantic concepts and order for image and sentence matching,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.170650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.790384Z digest=sha256:6c8d6935ced38fcfc830ee54ce49fe16e16f5f7c599864a4bf70a4dc3fa1a25f

Observation ee5e49a4-9c8a-4ed9-be2d-5070c4577c11 · outbound

This paper cites Stacked cross attention for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Stacked cross attention for image-text matching,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.156138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.794558Z digest=sha256:221962300bf98eb8b85d8a02acc190aed9fec14130bcf64911cb436c7a2297c3

Observation 8fd228ad-e85b-43c2-a565-2c2e4aa3d91a · outbound

This paper cites Cross- modal attention with semantic consistence for image–text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross- modal attention with semantic consistence for image–text matching,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.142074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.798448Z digest=sha256:7f7bd5d6777310215dd0344a8b5bac871012739ce491177f7f20d18f9734c990

Observation 38ae2995-86c3-4b14-93b2-da73ff3a6f7f · outbound

This paper cites Visual semantic reasoning for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Visual semantic reasoning for image-text matching,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.129587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.802192Z digest=sha256:290bfef973f2ab8fd9dd15e42af9a67caab9a8b1f2dcb91df0800f45da2ec094

Observation 91c20eb1-1199-455f-a128-3099ece39e98 · outbound

This paper cites Image-text embedding learning via visual and textual semantic reasoning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image-text embedding learning via visual and textual semantic reasoning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.115153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.805867Z digest=sha256:ae459aba507246f48bc323497fcbe29335fc63bd6bab925e51fdd51912b67341

Observation 1414cfaf-f4aa-449e-a1a5-f582b8bb01c7 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.809534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.809534Z digest=sha256:c4dcb7f7b9dd3663b33c04db902752d5f64a92de4bc269b999f96ca7503b97c3

Observation ecf46abb-c98e-4236-af91-6bb7d5159778 · outbound

This paper cites Lxmert: Learning cross-modality encoder representations from transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Lxmert: Learning cross-modality encoder representations from transformers,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.093662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.813277Z digest=sha256:e0105b175338cdf7076963c8b752e82f84c65acaf8d88ed2b850b5919b377196

Observation c1a80255-d52c-4e74-87fb-1c654dfcc07f · outbound

This paper cites Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.080079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.816997Z digest=sha256:d82b217603ba3d27d34b80ce47eb5876d665c84fbef0325e77f6cf89ac183e04

Observation 84f014d7-3af0-4263-b883-94ef99a55daa · outbound

This paper cites Learning the best pooling strategy for visual semantic embedding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning the best pooling strategy for visual semantic embedding,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.066838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.820850Z digest=sha256:c5bee39ccf6454ea449c1e47127a3f362d25ba4d1975ec5547bc43722d2f2d1c

Observation 0e90948c-19b9-4e50-8e10-f5988d1ed44f · outbound

This paper cites Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.824935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.824935Z digest=sha256:f77c5db972c8796c9356d4cc2f1a801dfba8aa254bee12fc16a0a5ff3307db4b

Observation bf64ed5d-685c-43e1-9734-415474ed2ece · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning transferable visual models from natural language supervision,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.829050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.829050Z digest=sha256:940eb8c7d569d9debefc66a1c2fba46ceff5f4a1b4aa45c2f059d843432d3c48

Observation e2c38548-073c-429f-a7d2-c76397f0bc7a · outbound

This paper cites Vista: Vision and scene text aggregation for cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vista: Vision and scene text aggregation for cross-modal retrieval,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.043211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.832696Z digest=sha256:4ff3c5031fddc291a9146574b8799e838c172296121cefeba9b49656a9530a5d

Observation d6f8ddff-9a5c-43b6-82b6-5bc90cca8ffb · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.027761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.836567Z digest=sha256:cc03d348125648b763db5aab03a0635e94d84a24e326f0a5f91ed8eba2f8d4c5

Observation 597fa9db-e019-4720-8555-7da321fee29a · outbound

This paper cites Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.013236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.840435Z digest=sha256:fb6c3b71253da3bdb0af4f24af8623436a732894a9f8f9c2c7e3f5d0a0592528

Observation 45903d78-eded-4900-a047-c4855e2cfac6 · outbound

This paper cites Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.999616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.844283Z digest=sha256:185b4629e3573a1952cb6edc9cf9131d562b1aebe8d9d441447a87210a0b29ca

Observation 9b3afb81-747a-406f-858f-61a9b7dd3591 · outbound

This paper cites Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.983362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.848175Z digest=sha256:a585aae7cd4d4095b0dcb279fe4158a2d468af4e7bb1eeb047409575aa0997a5

Observation 87a56034-0031-4957-9722-3c6a38a28750 · outbound

This paper cites Parameter-efficient transfer learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Parameter-efficient transfer learning for remote sensing image-text retrieval,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.852135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.852135Z digest=sha256:6cdab64ea7c09e7e1543dedef9909efc329569fec1ecb92ad8ffe25e34489cf2

Observation 8e1b2a61-fe9a-4da8-bcee-5a7041d2c5ce · outbound

This paper cites Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.856091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.856091Z digest=sha256:4e248718da2ee13a02fdf0a6839541dde7573d5280b554df00a047d5dec3e7be

Observation 1ac19980-e07b-4711-8c32-64bc5507a56c · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.951457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.859812Z digest=sha256:1063219bc3285433f5cdb04f61d82c3086c6cc88721aeb28db9950919908c025

Observation 852a2f44-d06f-416e-843f-3fb4879354b6 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Momentum contrast for unsupervised visual representation learning,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.937860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T15:05:57.863886Z digest=sha256:0d180aff5e3207cb13df37ccb7a2fbcc20c4ff2939372c9efd4e7990e9184381

Pith citing papers

Observation be331e05-aba8-4882-8bdd-c298a7fd228c · inbound

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing cites this paper.

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.841047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T00:02:12.727566Z digest=sha256:35628cc01dfdc2955ad929265b8cf6f1c7dc6321d38d2413d9303e8e5bd24e13