SCP: SCENE COMPLETION PRE-TRAINING FOR 3D OBJECT DETECTION

Shan, Y.; Xia, Y.; Chen, Y.; Cremers, D.

doi:https://doi.org/10.5194/isprs-archives-XLVIII-1-W2-2023-41-2023

Articles | Volume XLVIII-1/W2-2023

https://doi.org/10.5194/isprs-archives-XLVIII-1-W2-2023-41-2023

© Author(s) 2023. This work is distributed under
the Creative Commons Attribution 4.0 License.

https://doi.org/10.5194/isprs-archives-XLVIII-1-W2-2023-41-2023

© Author(s) 2023. This work is distributed under
the Creative Commons Attribution 4.0 License.

Articles | Volume XLVIII-1/W2-2023

13 Dec 2023

| 13 Dec 2023

SCP: SCENE COMPLETION PRE-TRAINING FOR 3D OBJECT DETECTION

Y. Shan, Y. Xia, Y. Chen, and D. Cremers

Keywords: LiDAR Point Clouds, 3D Object Detection, Pre-training, Scene Completion, Autonomous Driving

Abstract. 3D object detection using LiDAR point clouds is a fundamental task in the fields of computer vision, robotics, and autonomous driving. However, existing 3D detectors heavily rely on annotated datasets, which are both time-consuming and prone to errors during the process of labeling 3D bounding boxes. In this paper, we propose a Scene Completion Pre-training (SCP) method to enhance the performance of 3D object detectors with less labeled data. SCP offers three key advantages: (1) Improved initialization of the point cloud model. By completing the scene point clouds, SCP effectively captures the spatial and semantic relationships among objects within urban environments. (2) Elimination of the need for additional datasets. SCP serves as a valuable auxiliary network that does not impose any additional efforts or data requirements on the 3D detectors. (3) Reduction of the amount of labeled data for detection. With the help of SCP, the existing state-of-the-art 3D detectors can achieve comparable performance while only relying on 20% labeled data.

SCP: SCENE COMPLETION PRE-TRAINING FOR 3D OBJECT DETECTION

Useful Links

Useful External Links

Our Contact