Pith. sign in

REVIEW 2 cited by

Dynamic View Synthesis from Dynamic Monocular Video

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2105.06468 v1 pith:RL5OEW25 submitted 2021-05-13 cs.CV

classification cs.CV
keywords dynamicvideoimplicitinputmonocularnerfresultsscene
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present an algorithm for generating novel views at arbitrary viewpoints and any input time step given a monocular video of a dynamic scene. Our work builds upon recent advances in neural implicit representation and uses continuous and differentiable functions for modeling the time-varying structure and the appearance of the scene. We jointly train a time-invariant static NeRF and a time-varying dynamic NeRF, and learn how to blend the results in an unsupervised manner. However, learning this implicit function from a single video is highly ill-posed (with infinitely many solutions that match the input video). To resolve the ambiguity, we introduce regularization losses to encourage a more physically plausible solution. We show extensive quantitative and qualitative results of dynamic view synthesis from casually captured videos.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Hallo4D uses vision-language models to detect and correct spatial and temporal mistakes in AI-generated 3D and 4D content, improving consistency without retraining the base generators.

  2. Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos

    cs.CV 2025-06 conditional novelty 6.0 of 10

    A layered neural radiance field fused with 2D motion masks and refined at test time beats both the 2D motion segmentation baseline and previous 3D methods on dynamic object segmentation in egocentric video.

Pith tools