REVIEW 5 cited by
The QXS-SAROPT Dataset for Deep Learning in SAR-Optical Data Fusion
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Deep learning techniques have made an increasing impact on the field of remote sensing. However, deep neural networks based fusion of multimodal data from different remote sensors with heterogenous characteristics has not been fully explored, due to the lack of availability of big amounts of perfectly aligned multi-sensor image data with diverse scenes of high resolutions, especially for synthetic aperture radar (SAR) data and optical imagery. To promote the development of deep learning based SAR-optical fusion approaches, we release the QXS-SAROPT dataset, which contains 20,000 pairs of SAR-optical image patches. We obtain the SAR patches from SAR satellite GaoFen-3 images and the optical patches from Google Earth images. These images cover three port cities: San Diego, Shanghai and Qingdao. Here, we present a detailed introduction of the construction of the dataset, and show its two representative exemplary applications, namely SAR-optical image matching and SAR ship detection boosted by cross-modal information from optical images. As a large open SAR-optical dataset with multiple scenes of a high resolution, we believe QXS-SAROPT will be of potential value for further research in SAR-optical data fusion technology based on deep learning.
Forward citations
Cited by 5 Pith papers
-
LoRetta: A Foundation Model and Extensive Dataset for Global-Scale Remote Sensing Dense Image Matching
On the new LEVIR-GM benchmark, LoRetta's matchability-aware affine localization plus guided dense registration achieves AUC 83.3%, improving on RoMa v2 by 1.6 points while cutting inference time by 47.8%.
-
GeoCore-9B: Towards Geo-Aware Generative Foundation Models in Earth Observation
A 9B-parameter flow-matching diffusion transformer trained from scratch on satellite data, conditioned on text and geospatial metadata, sets new state-of-the-art results on several Earth observation generation and tra...
-
Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images
SegEarth-OV performs annotation-free open-vocabulary segmentation of remote-sensing images by upsampling CLIP features, removing global bias, and distilling optical knowledge into a SAR encoder.
-
Text-Guided Coarse-to-Fine Fusion Network for Robust Remote Sensing Visual Question Answering
An optical-SAR visual question answering dataset with 6,008 image pairs and 1,036,694 questions is introduced, together with a text-guided fusion network that outperforms baselines on that dataset.
-
C-DiffSET: Leveraging Latent Diffusion for SAR-to-EO Image Translation with Confidence-Guided Reliable Object Generation
Fine-tuning Stable Diffusion's U-Net with a confidence-weighted noise loss translates SAR to optical imagery with large FID and LPIPS gains over GAN and diffusion baselines on three datasets.
Discussion (0). Continue with ORCID to comment.