Pith. sign in

REVIEW 3 cited by

Depth-supervised NeRF: Fewer Views and Faster Training for Free

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2107.02791 v3 pith:S377SIMR submitted 2021-07-06 cs.CV cs.GRcs.LG

classification cs.CVcs.GRcs.LG
keywords depthnerftrainingds-nerfgivenlossradiancesupervision
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A commonly observed failure mode of Neural Radiance Field (NeRF) is fitting incorrect geometries when given an insufficient number of input views. One potential reason is that standard volumetric rendering does not enforce the constraint that most of a scene's geometry consist of empty space and opaque surfaces. We formalize the above assumption through DS-NeRF (Depth-supervised Neural Radiance Fields), a loss for learning radiance fields that takes advantage of readily-available depth supervision. We leverage the fact that current NeRF pipelines require images with known camera poses that are typically estimated by running structure-from-motion (SFM). Crucially, SFM also produces sparse 3D points that can be used as "free" depth supervision during training: we add a loss to encourage the distribution of a ray's terminating depth matches a given 3D keypoint, incorporating depth uncertainty. DS-NeRF can render better images given fewer training views while training 2-3x faster. Further, we show that our loss is compatible with other recently proposed NeRF methods, demonstrating that depth is a cheap and easily digestible supervisory signal. And finally, we find that DS-NeRF can support other types of depth supervision such as scanned depth sensors and RGB-D reconstruction outputs.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion

    cs.CV 2025-01 conditional novelty 6.0 of 10

    MVGD jointly generates novel-view images and scale-consistent depth maps with a pixel-level diffusion model, reporting state-of-the-art scores on several view synthesis and depth benchmarks.

  2. Reliability-Aware Monocular Depth Supervision for Sparse-View Neural Reconstruction

    cs.CV 2026-06 conditional novelty 4.0 of 10

    Masked monocular depth supervision improves Splatfacto PSNR and RMSE on sparse KITTI views by selecting low-photometric-error regions, while Mip-NeRF-360 gains little and object-centric scenes trade geometry for worse...

  3. Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning

    cs.RO 2025-08 conditional novelty 4.0 of 10

    The thesis demonstrates that combining implicit 3D scene representations with LLM-based reasoning, using text as an interface, yields strong performance on robotic perception and spatial language tasks.

Pith tools