Refined Voting and Scene Feature Fusion for 3D Object Detection in Point Clouds

Hang Yu; Jinhe Su; Yingchao Piao; Guorong Cai; Yangbin Lin; Niansheng Liu; Weiquan Liu

doi:10.1155/2022/3023934

Refined Voting and Scene Feature Fusion for 3D Object Detection in Point Clouds

Comput Intell Neurosci. 2022 Dec 29:2022:3023934. doi: 10.1155/2022/3023934. eCollection 2022.

Authors

Hang Yu¹, Jinhe Su¹, Yingchao Piao², Guorong Cai¹, Yangbin Lin¹, Niansheng Liu¹, Weiquan Liu³

Affiliations

¹ The School of Computer Engineering, Jimei University, Xiamen 361021, China.
² Computer Network Information Center, Chinese Academy of Sciences, Beijing, China.
³ Fujian Key Laboratory of Sensing and Computing for Smart Cities, School of Informatics, Xiamen University, Xiamen 361005, China.

Abstract

An essential task for 3D visual world understanding is 3D object detection in lidar point clouds. To predict directly bounding box parameters from point clouds, existing voting-based methods use Hough voting to obtain the centroid of each object. However, it may be difficult for the inaccurately voted centers to regress boxes accurately, leading to the generation of redundant bounding boxes. For objects in indoor scenes, there are several co-occurrence patterns for objects in indoor scenes. Concurrently, semantic relations between object layouts and scenes can be used as prior context to guide object detection. We propose a simple, yet effective network, RSFF-Net, which adds refined voting and scene feature fusion for indoor 3D object detection. The RSFF-Net consists of three modules: geometric function, refined voting, and scene constraint. First, a geometric function module is used to capture the geometric features of the nearest object of the voted points. Then, the coarse votes are revoted by a refined voting module, which is based on the fused feature between the coarse votes and geometric features. Finally, a scene constraint module is used to add the association information between candidate objects and scenes. RSFF-Net achieves competitive results on indoor 3D object detection benchmarks: ScanNet V2 and SUN RGB-D.

MeSH terms

Benchmarking*
Semantics*