Sequence training of multiple deep neural networks for better performance and faster training speed

Pan Zhou,Lirong Dai,Hui Jiang

doi:10.1109/icassp.2014.6854680

Abstract

Recently, sequence level discriminative training methods have been proposed to fine-tune deep neural networks (DNN) after the framelevel cross entropy (CE) training to further improve recognition performance of DNNs. In our previous work, we have proposed a new cluster-based multiple DNNs structure and its parallel training algorithm based on the frame-level cross entropy criterion, which can significantly expedite CE training with multiple GPUs. In this paper, we extend to full sequence training for the multiple DNNs structure for better performance and meanwhile we also consider a partial parallel implementation of sequence training using multiple GPUs for faster training speed. In this work, it is shown that sequence training can be easily extended to multiple DNNs by slightly modifying error signals in output layer. Many implementation steps in sequence training of multiple DNNs can still be parallelized across multiple GPUs for better efficiency. Experiments on the Switchboard task have shown that both frame-level CE training and sequence training of multiple DNNs can lead to massive training speedup with little degradation in recognition performance. Comparing with the state-of-the-art DNN, 4-cluster multiple DNNs model with similar size can achieve more than 7 times faster in CE training and about 1.5 times faster in sequence training when using 4 GPUs.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Sequence training of multiple deep neural networks for better performance and faster training speed

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

FAST: DNN Training Under Variable Precision Block Floating Point with Stochastic Rounding
Sai Qian Zhang ... H T Kung
-
Sai Qian Zhang, et. al.Sai Qian Zhang ... H T Kung
01 Apr 2022
01 Apr 2022

Neuroevolution in Deep Neural Networks: Current Trends and Future Challenges
Edgar Galvan ... Peter Mooney
IEEE Transactions on Artificial Intelligence | VOL. 2
Edgar Galvan, et. al.Edgar Galvan ... Peter Mooney
04 May 2021
IEEE Transactions on Artificial Intelligence | VOL. 2

State-Clustering Based Multiple Deep Neural Networks Modeling Approach for Speech Recognition
Pan Zhou ... Qing-Feng Liu
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 23
Pan Zhou, et. al.Pan Zhou ... Qing-Feng Liu
01 Apr 2015
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 23

A Framework for Distributed Deep Neural Network Training with Heterogeneous Computing Platforms
Bontak Gu ... Arslan Munir
-
Bontak Gu, et. al.Bontak Gu ... Arslan Munir
01 Dec 2019
01 Dec 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Sequence training of multiple deep neural networks for better performance and faster training speed

Abstract

Talk to us

Similar Papers