REVIEW 4 cited by
Learning Independent Instance Maps for Crowd Localization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Accurately locating each head's position in the crowd scenes is a crucial task in the field of crowd analysis. However, traditional density-based methods only predict coarse prediction, and segmentation/detection-based methods cannot handle extremely dense scenes and large-range scale-variations crowds. To this end, we propose an end-to-end and straightforward framework for crowd localization, named Independent Instance Map segmentation (IIM). Different from density maps and boxes regression, each instance in IIM is non-overlapped. By segmenting crowds into independent connected components, the positions and the crowd counts (the centers and the number of components, respectively) are obtained. Furthermore, to improve the segmentation quality for different density regions, we present a differentiable Binarization Module (BM) to output structured instance maps. BM brings two advantages into localization models: 1) adaptively learn a threshold map for different images to detect each instance more accurately; 2) directly train the model using loss on binary predictions and labels. Extensive experiments verify the proposed method is effective and outperforms the-state-of-the-art methods on the five popular crowd datasets. Significantly, IIM improves F1-measure by 10.4% on the NWPU-Crowd Localization task. The source code and pre-trained models will be released at https://github.com/taohan10200/IIM.
Forward citations
Cited by 4 Pith papers
-
Point-to-Region Loss for Semi-Supervised Point-Based Crowd Counting
A point-to-region loss that propagates pseudo-label confidence to background pixels makes semi-supervised point-based crowd counting work and beats prior methods on ShTech, UCF-QNRF, and JHU++.
-
Crowd Detection Using Very-Fine-Resolution Satellite Imagery
A new dataset and point-based network show that individual people in 0.3 m satellite imagery can be detected with about 66% F1-score, a new capability for large-scale crowd analysis.
-
Transformer-Based Dual-Optical Attention Fusion Crowd Head Point Counting and Localization Network
TAPNet fuses RGB and thermal imagery using attention and feature-decomposition modules and reports improved crowd counting and localization on two UAV datasets.
-
Text-guided Zero-Shot Object Localization
The authors introduce ZSOLNet, a CLIP-based text-guided zero-shot object localization framework with a text self-similarity matching module, evaluated on FSC-147, CARPK, and ShanghaiTech.
Discussion (0). Continue with ORCID to comment.