REVIEW 2 cited by
FDA: Fourier Domain Adaptation for Semantic Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We describe a simple method for unsupervised domain adaptation, whereby the discrepancy between the source and target distributions is reduced by swapping the low-frequency spectrum of one with the other. We illustrate the method in semantic segmentation, where densely annotated images are aplenty in one domain (synthetic data), but difficult to obtain in another (real images). Current state-of-the-art methods are complex, some requiring adversarial optimization to render the backbone of a neural network invariant to the discrete domain selection variable. Our method does not require any training to perform the domain alignment, just a simple Fourier Transform and its inverse. Despite its simplicity, it achieves state-of-the-art performance in the current benchmarks, when integrated into a relatively standard semantic segmentation model. Our results indicate that even simple procedures can discount nuisance variability in the data that more sophisticated methods struggle to learn away.
Forward citations
Cited by 2 Pith papers
-
MORDA: A Synthetic Dataset to Facilitate Adaptation of Object Detectors to Unseen Real-target Domain While Preserving Performance on Real-source Domain
MORDA, a synthetic dataset of South Korean digital twins with nuScenes-compatible sensors and labels, improves 2D/3D object detection on the unseen AI-Hub South Korea dataset when added to nuScenes training, while pre...
-
From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring
A ViT+FFT-ReLU cascade for image deblurring is reported as state-of-the-art, but its PSNR/SSIM gains over the ViT alone are negligible (0.00–0.03 dB) and the method's two-stage interface is never specified.
Discussion (0). Continue with ORCID to comment.