Transfer of Learning from Vision to Touch: A Hybrid Deep Convolutional Neural Network for Visuo-Tactile 3D Object Recognition

被引：8

作者：

Rouhafzay, Ghazal ^{[1
]}

Cretu, Ana-Maria ^{[2
]}

Payeur, Pierre ^{[1
]}

机构：

[1] Univ Ottawa, Sch Elect Engn & Comp Sci, Ottawa, ON K1N 6N5, Canada

[2] Univ Quebec Outaouais, Dept Comp Sci & Engn, Gatineau, PQ J8X 3X7, Canada

来源：

SENSORS | 2021年 / 21卷 / 01期

基金：

加拿大自然科学与工程研究理事会;

关键词：

3D object recognition; transfer learning; machine intelligence; convolutional neural networks; tactile sensors; force-sensing resistor; Barrett Hand;

D O I：

10.3390/s21010113

中图分类号：

O65 [分析化学];

学科分类号：

070302 ; 081704 ;

摘要：

Transfer of learning or leveraging a pre-trained network and fine-tuning it to perform new tasks has been successfully applied in a variety of machine intelligence fields, including computer vision, natural language processing and audio/speech recognition. Drawing inspiration from neuroscience research that suggests that both visual and tactile stimuli rouse similar neural networks in the human brain, in this work, we explore the idea of transferring learning from vision to touch in the context of 3D object recognition. In particular, deep convolutional neural networks (CNN) pre-trained on visual images are adapted and evaluated for the classification of tactile data sets. To do so, we ran experiments with five different pre-trained CNN architectures and on five different datasets acquired with different technologies of tactile sensors including BathTip, Gelsight, force-sensing resistor (FSR) array, a high-resolution virtual FSR sensor, and tactile sensors on the Barrett robotic hand. The results obtained confirm the transferability of learning from vision to touch to interpret 3D models. Due to its higher resolution, tactile data from optical tactile sensors was demonstrated to achieve higher classification rates based on visual features compared to other technologies relying on pressure measurements. Further analysis of the weight updates in the convolutional layer is performed to measure the similarity between visual and tactile features for each technology of tactile sensing. Comparing the weight updates in different convolutional layers suggests that by updating a few convolutional layers of a pre-trained CNN on visual data, it can be efficiently used to classify tactile data. Accordingly, we propose a hybrid architecture performing both visual and tactile 3D object recognition with a MobileNetV2 backbone. MobileNetV2 is chosen due to its smaller size and thus its capability to be implemented on mobile devices, such that the network can classify both visual and tactile data. An accuracy of 100% for visual and 77.63% for tactile data are achieved by the proposed architecture.

引用

页码：1 / 15

页数：15

共 50 条

[41] Biologically Inspired Vision and Touch Sensing to Optimize 3D Object Representation and Recognition
Rouhafzay, Ghazal
Cretu, Ana-Maria
Payeur, Pierre
IEEE INSTRUMENTATION & MEASUREMENT MAGAZINE, 2021, 24 (03) : 85 - 90
[42] Learning Action Images Using Deep Convolutional Neural Networks For 3D Action Recognition
Thien Huynh-The
Hua, Cam-Hao
Kim, Dong-Seong
2019 IEEE SENSORS APPLICATIONS SYMPOSIUM (SAS), 2019,
[43] DRCNN: Dynamic Routing Convolutional Neural Network for Multi-View 3D Object Recognition
Sun, Kai
Zhang, Jiangshe
Liu, Junmin
Yu, Ruixuan
Song, Zengjie
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2021, 30 : 868 - 877
[44] Deep Learning of Volumetric Representation for 3D Object Recognition
Liu, Hongsen
Cong, Yang
Tang, Yandong
2017 32ND YOUTH ACADEMIC ANNUAL CONFERENCE OF CHINESE ASSOCIATION OF AUTOMATION (YAC), 2017, : 663 - 668
[45] Drcnn: Dynamic routing convolutional neural network for multi-view 3d object recognition
Sun, Kai
Zhang, Jiangshe
Liu, Junmin
Yu, Ruixuan
Song, Zengjie
IEEE Transactions on Image Processing, 2021, 30 : 868 - 877
[46] CURVATURE AUGMENTED DEEP LEARNING FOR 3D OBJECT RECOGNITION
Braeger, Sarah
Foroosh, Hassan
2018 25TH IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2018, : 3648 - 3652
[47] High-Performance Object Recognition by Employing a Transfer Learned Deep Convolutional Neural Network
Hasan, Md Mehedi
Srizon, Azmain Yakin
Abu Sayeed
Hasan, Md Al Mehedi
PROCEEDINGS OF 2020 11TH INTERNATIONAL CONFERENCE ON ELECTRICAL AND COMPUTER ENGINEERING (ICECE), 2020, : 250 - 253
[48] 3D palmprint recognition using unsupervised convolutional deep learning network and SVM classifier
Chaa, Mourad
Akhtar, Zahid
Attia, Abdelouahab
IET IMAGE PROCESSING, 2019, 13 (05) : 736 - 745
[49] LPI Radar Waveform Recognition Based on Deep Convolutional Neural Network Transfer Learning
Guo, Qiang
Yu, Xin
Ruan, Guoqing
SYMMETRY-BASEL, 2019, 11 (04):
[50] A Deep Learning Framework Using Convolutional Neural Network for Multi-class Object Recognition
Hayat, Shaukat
She Kun
Zuo Tengtao
Yue Yu
Tu, Tianyi
Du, Yantong
2018 IEEE 3RD INTERNATIONAL CONFERENCE ON IMAGE, VISION AND COMPUTING (ICIVC), 2018, : 194 - 198

← 1 2 3 4 5 →