Improved Depression Recognition Using Attention and Multitask Learning of Gender Recognition

Yang Liu,Daimin Shi,Xiaoyong Lu,Tao Pan,Jingyi Yuan,Haizhen An

doi:10.1109/ialp54817.2021.9675220

Abstract

College students suffer from depression due to factors such as education and graduation, and this phenomenon is increasing, but there is less research in this area. We studied depressive tendencies among Chinese students and used machine learning methods to detect depressive tendencies. The paper presents a Multi-Head Attention deep learning network for Speech Depression Recognition (SDR) using the Mel-frequency cepstral coefficient (MFCC) features as the input. The multi-head attention along with the convolutional neural network and the bidirectional long short-term memory network (CNN-BLSTM) embedding jointly attends to information from different representations of the same MFCC input sequence. The CNN-LSTM embedding helps in attending to the dominant depression features by identifying positions of the features in the sequence. In addition to Multi-Head Attention and CNN-LSTM embedding, we apply multi-task learning with gender recognition as an auxiliary task. The auxiliary task helps in learning the gender-specific features that influence the depression characteristics in speech and results in improved accuracy of Speech Depression Recognition, the primary task. We conducted all our experiments on Depression dataset. We can achieve an overall F1sorce of 91 % and average class accuracy of 92%, on SDR for depression classes.

Full Text