REVIEW 3 cited by
DSFD: Dual Shot Face Detector
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper, we propose a novel face detection network with three novel contributions that address three key aspects of face detection, including better feature learning, progressive loss design and anchor assign based data augmentation, respectively. First, we propose a Feature Enhance Module (FEM) for enhancing the original feature maps to extend the single shot detector to dual shot detector. Second, we adopt Progressive Anchor Loss (PAL) computed by two different sets of anchors to effectively facilitate the features. Third, we use an Improved Anchor Matching (IAM) by integrating novel anchor assign strategy into data augmentation to provide better initialization for the regressor. Since these techniques are all related to the two-stream design, we name the proposed network as Dual Shot Face Detector (DSFD). Extensive experiments on popular benchmarks, WIDER FACE and FDDB, demonstrate the superiority of DSFD over the state-of-the-art face detectors.
Forward citations
Cited by 3 Pith papers
-
Residual Objectness for Imbalance Reduction
Residual Objectness replaces hand-crafted sampling and reweighting with cascaded learned objectness refinements, improving RetinaNet, YOLOv3, and Faster R-CNN by 1.1 to 1.3 AP on COCO.
-
Relational Programming with Foundation Models
Vieira extends the Scallop relational engine with a foreign interface that lets foundation models act as probabilistic relations, enabling neuro-symbolic programs across nine tasks.
-
AFP-Net: Realtime Anchor-Free Polyp Detection in Colonoscopy
AFP-Net, an anchor-free polyp detector with a context enhancement module and cosine ground-truth projection, achieves 99.36% precision and 96.44% recall on CVC-Clinic, and 52.6 FPS.
Discussion (0). Continue with ORCID to comment.