REVIEW 3 cited by
Mixture of Counting CNNs: Adaptive Integration of CNNs Specialized to Specific Appearance for Crowd Counting
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
This paper proposes a crowd counting method. Crowd counting is difficult because of large appearance changes of a target which caused by density and scale changes. Conventional crowd counting methods generally utilize one predictor (e,g., regression and multi-class classifier). However, such only one predictor can not count targets with large appearance changes well. In this paper, we propose to predict the number of targets using multiple CNNs specialized to a specific appearance, and those CNNs are adaptively selected according to the appearance of a test image. By integrating the selected CNNs, the proposed method has the robustness to large appearance changes. In experiments, we confirm that the proposed method can count crowd with lower counting error than a CNN and integration of CNNs with fixed weights. Moreover, we confirm that each predictor automatically specialized to a specific appearance.
Forward citations
Cited by 3 Pith papers
-
Count2Density: Crowd Density Estimation without Location-level Annotations
Count2Density maps per-image crowd counts into spatial density maps via an EMA historical bank and hypergeometric sampling, eliminating the need for point annotations.
-
Attend To Count: Crowd Counting with Adaptive Capacity Multi-scale CNNs
A three-part CNN with a count attention mechanism routes dense and sparse image regions to networks of different capacities and reports state-of-the-art counting errors on five benchmarks.
-
Crowd Scene Analysis using Deep Learning Techniques
The paper describes a 5-column M-CNN with rotation self-supervision and Sinkhorn distribution matching for crowd counting and a VGG19-LSTM with dense residual blocks for violence detection, but the evidence does not s...
Discussion (0). Continue with ORCID to comment.