Pith. sign in

REVIEW 1 cited by

ActiveStereoNet: End-to-End Self-Supervised Learning for Active Stereo Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1807.06009 v1 pith:RFAFVLGR submitted 2018-07-16 cs.CV

ActiveStereoNet: End-to-End Self-Supervised Learning for Active Stereo Systems

classification cs.CV
keywords lossactiveactivestereonetaggregationcostdepthend-to-endlearning
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

In this paper we present ActiveStereoNet, the first deep learning solution for active stereo systems. Due to the lack of ground truth, our method is fully self-supervised, yet it produces precise depth with a subpixel precision of $1/30th$ of a pixel; it does not suffer from the common over-smoothing issues; it preserves the edges; and it explicitly handles occlusions. We introduce a novel reconstruction loss that is more robust to noise and texture-less patches, and is invariant to illumination changes. The proposed loss is optimized using a window-based cost aggregation with an adaptive support weight scheme. This cost aggregation is edge-preserving and smooths the loss function, which is key to allow the network to reach compelling results. Finally we show how the task of predicting invalid regions, such as occlusions, can be trained end-to-end without ground-truth. This component is crucial to reduce blur and particularly improves predictions along depth discontinuities. Extensive quantitatively and qualitatively evaluations on real and synthetic data demonstrate state of the art results in many challenging scenes.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. NSL-SLAM: High-Fidelity Neural Structured-Light Depth for Practical SLAM and Reconstruction

    cs.CV 2026-07 conditional novelty 5.0

    Injecting frozen monocular depth features into neural structured-light decoding cuts Replica-SL depth RMSE ~35% vs NSL and, with depth-centric GICP+sparse anchors+light BA, yields the most stable real D435 SLAM among ...