Point2SpatialCapsule: Aggregating Features and Spatial Relationships of Local Regions on Point Clouds Using Spatial-Aware Capsules

被引:31
作者
Wen, Xin [1 ]
Han, Zhizhong [2 ]
Liu, Xinhai [1 ]
Liu, Yu-Shen [3 ]
机构
[1] Tsinghua Univ, Sch Software, Beijing 100084, Peoples R China
[2] Univ Maryland, Dept Comp Sci, College Pk, MD 20737 USA
[3] Tsinghua Univ, BNRist, Sch Software, Beijing 100084, Peoples R China
关键词
Three-dimensional displays; Feature extraction; Shape; Routing; Aggregates; Machine learning; Spatial resolution; Point cloud; shape representation; feature aggregation; spatial relationships; capsule network; SHAPE; REPRESENTATION; PREDICTION; NETWORK; VIEW;
D O I
10.1109/TIP.2020.3019925
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Learning discriminative shape representation directly on point clouds is still challenging in 3D shape analysis and understanding. Recent studies usually involve three steps: first splitting a point cloud into some local regions, then extracting the corresponding feature of each local region, and finally aggregating all individual local region features into a global feature as shape representation using simple max-pooling. However, such pooling-based feature aggregation methods do not adequately take the spatial relationships (e.g. the relative locations to other regions) between local regions into account, which greatly limits the ability to learn discriminative shape representation. To address this issue, we propose a novel deep learning network, named Point2SpatialCapsule, for aggregating features and spatial relationships of local regions on point clouds, which aims to learn more discriminative shape representation. Compared with the traditional max-pooling based feature aggregation networks, Point2SpatialCapsule can explicitly learn not only geometric features of local regions but also the spatial relationships among them. Point2SpatialCapsule consists of two main modules. To resolve the disorder problem of local regions, the first module, named geometric feature aggregation, is designed to aggregate the local region features into the learnable cluster centers, which explicitly encodes the spatial locations from the original 3D space. The second module, named spatial relationship aggregation, is proposed for further aggregating the clustered features and the spatial relationships among them in the feature space using the spatial-aware capsules developed in this article. Compared to the previous capsule network based methods, the feature routing on the spatial-aware capsules can learn more discriminative spatial relationships among local regions for point clouds, which establishes a direct mapping between log priors and the spatial locations through feature clusters. Experimental results demonstrate that Point2SpatialCapsule outperforms the state-of-the-art methods in the 3D shape classification, retrieval and segmentation tasks under the well-known ModelNet and ShapeNet datasets.
引用
收藏
页码:8855 / 8869
页数:15
相关论文
共 84 条
[1]  
[Anonymous], 2017, ARXIV171110108
[2]  
[Anonymous], 2018, ARXIV180404241
[3]  
Arandjelovic R, 2018, IEEE T PATTERN ANAL, V40, P1437, DOI [10.1109/CVPR.2016.572, 10.1109/TPAMI.2017.2711011]
[4]   GIFT: Towards Scalable 3D Shape Retrieval [J].
Bai, Song ;
Bai, Xiang ;
Zhou, Zhichao ;
Zhang, Zhaoxiang ;
Tian, Qi ;
Latecki, Longin Jan .
IEEE TRANSACTIONS ON MULTIMEDIA, 2017, 19 (06) :1257-1271
[5]   Deep Unsupervised Learning of 3D Point Clouds via Graph Topology Inference and Filtering [J].
Chen, Siheng ;
Duan, Chaojing ;
Yang, Yaoqing ;
Li, Duanshun ;
Feng, Chen ;
Tian, Dong .
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2020, 29 :3183-3198
[6]   3DCapsule: Extending the Capsule Architecture to Classify 3D Point Clouds [J].
Cheraghian, Ali ;
Petersson, Lars .
2019 IEEE WINTER CONFERENCE ON APPLICATIONS OF COMPUTER VISION (WACV), 2019, :1194-1202
[7]   Shape Completion using 3D-Encoder-Predictor CNNs and Shape Synthesis [J].
Dai, Angela ;
Qi, Charles Ruizhongtai ;
Niessner, Matthias .
30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :6545-6554
[8]   ScanNet: Richly-annotated 3D Reconstructions of Indoor Scenes [J].
Dai, Angela ;
Chang, Angel X. ;
Savva, Manolis ;
Halber, Maciej ;
Funkhouser, Thomas ;
Niessner, Matthias .
30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :2432-2443
[9]   Enhancement of critical current density in CaKFe4As4 single crystals through 3 MeV proton irradiation [J].
Haberkorn, N. ;
Xu, M. ;
Meier, W. R. ;
Suarez, S. ;
Bud'ko, S. L. ;
Canfield, P. C. .
SUPERCONDUCTOR SCIENCE & TECHNOLOGY, 2020, 33 (02)
[10]  
Han W., 2020, P 33 AAAI C ART INT