Pith. sign in

REVIEW 2 cited by

Point2Graph: An End-to-end Point Cloud-based 3D Open-Vocabulary Scene Graph for Robot Navigation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.10350 v1 pith:FQXDNRUN submitted 2024-09-16 cs.RO cs.AIcs.CV

classification cs.ROcs.AIcs.CV
keywords open-vocabularyframeworkpointroomsceneclassificationdetectionend-to-end
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Current open-vocabulary scene graph generation algorithms highly rely on both 3D scene point cloud data and posed RGB-D images and thus have limited applications in scenarios where RGB-D images or camera poses are not readily available. To solve this problem, we propose Point2Graph, a novel end-to-end point cloud-based 3D open-vocabulary scene graph generation framework in which the requirement of posed RGB-D image series is eliminated. This hierarchical framework contains room and object detection/segmentation and open-vocabulary classification. For the room layer, we leverage the advantage of merging the geometry-based border detection algorithm with the learning-based region detection to segment rooms and create a "Snap-Lookup" framework for open-vocabulary room classification. In addition, we create an end-to-end pipeline for the object layer to detect and classify 3D objects based solely on 3D point cloud data. Our evaluation results show that our framework can outperform the current state-of-the-art (SOTA) open-vocabulary object and room segmentation and classification algorithm on widely used real-scene datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. OpenGuide: Assistive Object Retrieval in Indoor Spaces for Individuals with Visual Impairments

    cs.RO 2025-09 conditional novelty 6.0 of 10

    OpenGuide combines vision-language value maps, frontier exploration, and POMDP planning to locate multiple objects in unfamiliar indoor spaces, reaching about 55% success in simulation and 54% in real-world trials.

  2. OpenTie: Open-vocabulary Sequential Rebar Tying System

    cs.RO 2025-08 reject novelty 4.0 of 10

    A claimed training-free rebar tying pipeline based on point clouds and open-vocabulary detection, but the reported evaluation is too vague to verify the claimed 90% success.

Pith tools