REVIEW 3 cited by
OpenRooms: An End-to-End Open Framework for Photorealistic Indoor Scene Datasets
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We propose a novel framework for creating large-scale photorealistic datasets of indoor scenes, with ground truth geometry, material, lighting and semantics. Our goal is to make the dataset creation process widely accessible, transforming scans into photorealistic datasets with high-quality ground truth for appearance, layout, semantic labels, high quality spatially-varying BRDF and complex lighting, including direct, indirect and visibility components. This enables important applications in inverse rendering, scene understanding and robotics. We show that deep networks trained on the proposed dataset achieve competitive performance for shape, material and lighting estimation on real images, enabling photorealistic augmented reality applications, such as object insertion and material editing. We also show our semantic labels may be used for segmentation and multi-task learning. Finally, we demonstrate that our framework may also be integrated with physics engines, to create virtual robotics environments with unique ground truth such as friction coefficients and correspondence to real scenes. The dataset and all the tools to create such datasets will be made publicly available.
Forward citations
Cited by 3 Pith papers
-
DNF-Intrinsic: Deterministic Noise-Free Diffusion for Indoor Inverse Rendering
A diffusion model fine-tuned with flow matching maps a single RGB indoor image directly to five scene properties, beating prior inverse rendering methods on InteriorVerse and real-world benchmarks.
-
UniRelight: Learning Joint Decomposition and Synthesis for Video Relighting
Jointly predicting albedo and relit appearance with one video-diffusion pass improves relighting fidelity and generalization over two-stage inverse-plus-forward pipelines.
-
DiffusionRenderer: Neural Inverse and Forward Rendering with Video Diffusion Models
A single video diffusion system both estimates scene properties from video and renders photorealistic images from those properties, enabling relighting, material editing, and object insertion.
Discussion (0). Continue with ORCID to comment.