Pith. sign in

REVIEW 3 cited by

Zero-shot point cloud segmentation by transferring geometric primitives

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.09923 v3 pith:EIEY4LSB submitted 2022-10-18 cs.CV

classification cs.CV
keywords geometricnovelprimitiveslanguageobjectspointfine-grainedcloud
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We investigate transductive zero-shot point cloud semantic segmentation, where the network is trained on seen objects and able to segment unseen objects. The 3D geometric elements are essential cues to imply a novel 3D object type. However, previous methods neglect the fine-grained relationship between the language and the 3D geometric elements. To this end, we propose a novel framework to learn the geometric primitives shared in seen and unseen categories' objects and employ a fine-grained alignment between language and the learned geometric primitives. Therefore, guided by language, the network recognizes the novel objects represented with geometric primitives. Specifically, we formulate a novel point visual representation, the similarity vector of the point's feature to the learnable prototypes, where the prototypes automatically encode geometric primitives via back-propagation. Besides, we propose a novel Unknown-aware InfoNCE Loss to fine-grained align the visual representation with language. Extensive experiments show that our method significantly outperforms other state-of-the-art methods in the harmonic mean-intersection-over-union (hIoU), with the improvement of 17.8\%, 30.4\%, 9.2\% and 7.9\% on S3DIS, ScanNet, SemanticKITTI and nuScenes datasets, respectively. Codes are available (https://github.com/runnanchen/Zero-Shot-Point-Cloud-Segmentation)

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

    cs.CV 2025-07 conditional novelty 6.0 of 10

    A large-scale 3D spatial reasoning segmentation benchmark with human-written queries that avoid object names shows current 3D vision-language models underperform.

  2. PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM

    cs.CV 2024-12 conditional novelty 6.0 of 10

    PanoSLAM is a Gaussian Splatting SLAM system that produces label-free 3D panoptic maps from RGB-D video by lifting and refining 2D panoptic predictions in 3D.

  3. OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies

    cs.CV 2024-12 conditional novelty 6.0 of 10

    This paper introduces a dataset and a cross-modal network that predicts renderable open-vocabulary semantic attributes for 3D Gaussian scenes, enabling segmentation of new scenes without scene-specific fine-tuning.

Pith tools