A novel privacy-preserving speech recognition framework using bidirectional LSTM

Qingren Wang,Chuankai Feng,Yan Xu,Victor S Sheng,Hong Zhong

doi:10.1186/s13677-020-00186-7

Qingren Wang, Chuankai Feng + Show 3 more

Open Access

https://doi.org/10.1186/s13677-020-00186-7

Copy DOI

Journal: Journal of Cloud Computing	Publication Date: Jul 8, 2020
Citations: 25	License type: open-access

Affiliation: Anhui University, Texas Tech University

Abstract

Utilizing speech as the transmission medium in Internet of things (IoTs) is an effective way to reduce latency while improving the efficiency of human-machine interaction. In the field of speech recognition, Recurrent Neural Network (RNN) has significant advantages to achieve accuracy improvement on speech recognition. However, some of RNN-based intelligence speech recognition applications are insufficient in the privacy-preserving of speech data, and others with privacy-preserving are time-consuming, especially about model training and speech recognition. Therefore, in this paper we propose a novel Privacy-preserving Speech Recognition framework using Bidirectional Long short-term memory neural network, namely PSRBL. On the one hand, PSRBL designs new functions to construct security activation functions by combing with an additive secret sharing protocol, namely a secure piecewise-linear Sigmoid and a secure piecewise-linear Tanh respectively, to achieve privacy-preserving of speech data during speech recognition process running on edge servers. On the other hand, in order to reduce the time spent on both the training and the recognition of the speech model while keeping high accuracy during speech recognition process, PSRBL first utilizes secure activation functions to refit original activation functions in the bidirectional Long Short-Term Memory neural network (LSTM), and then makes full use of the left and the right context information of speech data by employing bidirectional LSTM. Experiments conducted on the speech dataset TIMIT show that our framework PSRBL performs well. Specifically compared with the state-of-the-art ones, PSRBL significantly reduces the time consumption on both the training and the recognition of the speech model under the premise that PSRBL and the comparisons are consistent in the privacy-preserving of speech data.

Highlights

Utilizing speech as the transmission medium in Internet of things (IoTs) is an effective way to reduce latency while improving the efficiency of human-machine interactions
Studies have demonstrated that speech recognition applications based on Recurrent Neural Network (RNN) perform well in terms of improving accuracy of speech recognition [3]
The experimental results demonstrate that our PSRBL significantly reduces the time consumption in terms of the training and the recognition of models, and improves the response speed while preserving the privacy information of the speech data on the edge servers

Summary

Introduction

Utilizing speech as the transmission medium in Internet of things (IoTs) is an effective way to reduce latency while improving the efficiency of human-machine interactions. The data traffic of these explosively increasing terminal devices is transmitted to the cloud for processing, which will eventually exceed the cloud’s computing and storage capabilities. Most of RNN-based speech recognition applications are deployed on edge servers to alleviate challenges derived from computing-intensiveness and insufficient storage capabilities of the clouds [4, 5]. In the era of big data with the explosive growth of data volume, deploying a speech recognition application on edge servers is not an effective solution since edge computing suffers from capacity limitations. The edge-cloud computing paradigm offers a tradeoff between speech recognition applications’ requirements for computing resources and low latency, and improves the usage efficiency of the IoT devices [6]

Methods

Findings

Conclusion

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A novel privacy-preserving speech recognition framework using bidirectional LSTM

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Journal of Cloud Computing

Lead the way for us

Similar Papers

Bidirectional recurrent neural network language models for automatic speech recognition
Ebru Arisoy ... Abhinav Sethy
-
Ebru Arisoy, et. al.Ebru Arisoy ... Abhinav Sethy
01 Apr 2015
01 Apr 2015

Wind Power Forecasting Method Based on Bidirectional Long Short-Term Memory Neural Network and Error Correction
Wei Liu ... Yanping Kang
Electric Power Components and Systems | VOL. ahead-of-print
Wei Liu, et. al.Wei Liu ... Yanping Kang
08 Mar 2022
Electric Power Components and Systems | VOL. ahead-of-print

Lithium Battery State-of-Charge Estimation Based on a Bayesian Optimization Bidirectional Long Short-Term Memory Neural Network
Biao Yang ... Yuedong Zhan
Energies | VOL. 15
Biao Yang, et. al.Biao Yang ... Yuedong Zhan
25 Jun 2022
Energies | VOL. 15

Improved Protein Secondary Structure Prediction Using Bidirectional Long Short-Term Memory Neural Network and Bootstrap Aggregating
Wen-Wu Zeng ... Jun Hu
-
Wen-Wu Zeng, et. al.Wen-Wu Zeng ... Jun Hu
13 May 2022
13 May 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A novel privacy-preserving speech recognition framework using bidirectional LSTM

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Journal of Cloud Computing