REVIEW 12 cited by
SAGD: Boundary-Enhanced Segment Anything in 3D Gaussian via Gaussian Decomposition
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
3D Gaussian Splatting has emerged as an alternative 3D representation for novel view synthesis, benefiting from its high-quality rendering results and real-time rendering speed. However, the 3D Gaussians learned by 3D-GS have ambiguous structures without any geometry constraints. This inherent issue in 3D-GS leads to a rough boundary when segmenting individual objects. To remedy these problems, we propose SAGD, a conceptually simple yet effective boundary-enhanced segmentation pipeline for 3D-GS to improve segmentation accuracy while preserving segmentation speed. Specifically, we introduce a Gaussian Decomposition scheme, which ingeniously utilizes the special structure of 3D Gaussian, finds out, and then decomposes the boundary Gaussians. Moreover, to achieve fast interactive 3D segmentation, we introduce a novel training-free pipeline by lifting a 2D foundation model to 3D-GS. Extensive experiments demonstrate that our approach achieves high-quality 3D segmentation without rough boundary issues, which can be easily applied to other scene editing tasks.
Forward citations
Cited by 12 Pith papers
-
GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization
A training-free graph-cut method selects 3D objects from Gaussian splatting scenes using sparse user scribbles, reaching 92.2 mIoU on NVOS with three interaction views.
-
LabelGS: Label-Aware 3D Gaussian Splatting for 3D Scene Segmentation
LabelGS assigns 2D video-tracking labels to the most-contributing 3D Gaussians, with depth-based occlusion masking and a projection filter, reporting better mIoU/PSNR than Feature-3DGS with roughly 22x faster training.
-
DCHM: Depth-Consistent Human Modeling for Multiview Detection
DCHM uses superpixel-based Gaussian Splatting to make monocular depth estimates multiview-consistent, producing point clouds that yield state-of-the-art label-free pedestrian detection on Wildtrack, Terrace, and MultiviewX.
-
TexGS-VolVis: Expressive Scene Editing for Volume Visualization via Textured Gaussian Splatting
A textured Gaussian splatting framework enables flexible image- and text-driven style editing of volume visualizations with real-time rendering.
-
VolSegGS: Segmentation and Tracking in Dynamic Volumetric Scenes via Deformable 3D Gaussians
VolSegGS reconstructs dynamic volumetric scenes from rendered images with deformable 3D Gaussians and enables real-time interactive segmentation and tracking of regions over time.
-
NLI4VolVis: Natural Language Interaction for Volume Visualization via LLM Multi-Agents and Editable 3D Gaussian Splatting
NLI4VolVis integrates multi-agent large language models, editable 3D Gaussian splatting, and CLIP-based querying so users can explore, query, and edit volume visualizations through natural language.
-
BRUM: Robust 3D Vehicle Reconstruction from 360 Sparse Images
BRUM improves sparse-view 3D vehicle reconstruction by synthesizing extra training views from depth and poses, using DUSt3R for real cameras, and weighting the photometric loss by pixel reliability.
-
DSG-World: Learning a 3D Gaussian World Model from Dual State Videos
DSG-World builds two segmented 3D Gaussian fields from two scene states and trains them with mutual consistency, enabling novel-state simulation without inpainting or dense capture.
-
Tackling View-Dependent Semantics in 3D Language Gaussian Splatting
A new 3D language Gaussian Splatting method that clusters per-object multi-view CLIP features and reweights them to capture view-dependent semantics, improving direct 3D open-vocabulary segmentation.
-
Enhancing LLM Training via Spectral Clipping
SPECTRA improves LLM pretraining via post-clipping of update spectral norms and optional pre-clipping of gradient spikes, framed as Composite Frank-Wolfe regularization.
-
Leveraging 2D Priors and SDF Guidance for Dynamic Urban Scene Rendering
UGSDF achieves state-of-the-art novel-view rendering of dynamic urban objects without LiDAR or 3D motion annotations by jointly optimizing SDFs and 3D Gaussians under 2D depth and point-tracking priors.
-
The ALMA-QUARKS Survey: III. Clump-to-core fragmentation and search for high-mass starless cores
In 139 infrared-bright massive protoclusters, ALMA resolves 1562 cores whose separations are much smaller than the Jeans length, and finds only two candidate high-mass starless cores.
Discussion (0). Continue with ORCID to comment.