Transfer of Learning from Vision to Touch: A Hybrid Deep Convolutional Neural Network for Visuo-Tactile 3D Object Recognition.

Ghazal Rouhafzay,Pierre Payeur,Ana-Maria Cretu

doi:10.3390/s21010113

Ghazal Rouhafzay, Pierre Payeur + Show 1 more

Open Access

https://doi.org/10.3390/s21010113

Copy DOI

Abstract

Transfer of learning or leveraging a pre-trained network and fine-tuning it to perform new tasks has been successfully applied in a variety of machine intelligence fields, including computer vision, natural language processing and audio/speech recognition. Drawing inspiration from neuroscience research that suggests that both visual and tactile stimuli rouse similar neural networks in the human brain, in this work, we explore the idea of transferring learning from vision to touch in the context of 3D object recognition. In particular, deep convolutional neural networks (CNN) pre-trained on visual images are adapted and evaluated for the classification of tactile data sets. To do so, we ran experiments with five different pre-trained CNN architectures and on five different datasets acquired with different technologies of tactile sensors including BathTip, Gelsight, force-sensing resistor (FSR) array, a high-resolution virtual FSR sensor, and tactile sensors on the Barrett robotic hand. The results obtained confirm the transferability of learning from vision to touch to interpret 3D models. Due to its higher resolution, tactile data from optical tactile sensors was demonstrated to achieve higher classification rates based on visual features compared to other technologies relying on pressure measurements. Further analysis of the weight updates in the convolutional layer is performed to measure the similarity between visual and tactile features for each technology of tactile sensing. Comparing the weight updates in different convolutional layers suggests that by updating a few convolutional layers of a pre-trained CNN on visual data, it can be efficiently used to classify tactile data. Accordingly, we propose a hybrid architecture performing both visual and tactile 3D object recognition with a MobileNetV2 backbone. MobileNetV2 is chosen due to its smaller size and thus its capability to be implemented on mobile devices, such that the network can classify both visual and tactile data. An accuracy of 100% for visual and 77.63% for tactile data are achieved by the proposed architecture.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Sensors (Basel, Switzerland)	Publication Date: Dec 27, 2020
Citations: 10	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Transfer of Learning from Vision to Touch: A Hybrid Deep Convolutional Neural Network for Visuo-Tactile 3D Object Recognition.

Abstract

Talk to us

Similar Papers

More From: Sensors (Basel, Switzerland)

Lead the way for us

Similar Papers

DCNN for Tactile Sensory Data Classification based on Transfer Learning
Mohamad Alameh ... Maurizio Valle
-
Mohamad Alameh, et. al.Mohamad Alameh ... Maurizio Valle
01 Jul 2019
01 Jul 2019

Automated detection of COVID-19 using ensemble of transfer learning with deep convolutional neural network based on CT scans.
Parisa Gifani ... Majid Vafaeezadeh
International Journal of Computer Assisted Radiology and Surgery | VOL. 16
Parisa Gifani, et. al.Parisa Gifani ... Majid Vafaeezadeh
16 Nov 2020
International Journal of Computer Assisted Radiology and Surgery | VOL. 16

Performance analysis of pretrained convolutional neural network models for ophthalmological disease classification.
Busra Emir ... Ertugrul Colak
Arquivos brasileiros de oftalmologia | VOL. 87
Busra Emir, et. al.Busra Emir ... Ertugrul Colak
01 Jan 2023
Arquivos brasileiros de oftalmologia | VOL. 87

Bidirectional visual-tactile cross-modal generation using latent feature space flow model
Yu Fang ... Jie Zhao
Neural Networks | VOL. 172
Yu Fang, et. al.Yu Fang ... Jie Zhao
27 Dec 2023
Neural Networks | VOL. 172

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Transfer of Learning from Vision to Touch: A Hybrid Deep Convolutional Neural Network for Visuo-Tactile 3D Object Recognition.

Abstract

Talk to us

Similar Papers

More From: Sensors (Basel, Switzerland)