Audio based depression detection using Convolutional Autoencoder. (1st March 2022)
- Record Type:
- Journal Article
- Title:
- Audio based depression detection using Convolutional Autoencoder. (1st March 2022)
- Main Title:
- Audio based depression detection using Convolutional Autoencoder
- Authors:
- Sardari, Sara
Nakisa, Bahareh
Rastgoo, Mohammed Naim
Eklund, Peter - Abstract:
- Highlights: A novel audio-based depression detection system using Convolutional Autoencoder. Convolutional Autoencoder for extracting highly correlated and compact feature set. Thorough experimental study based on a real-world depression detection dataset. Complete comparison of proposed feature extraction method with other techniques. Abstract: Depression is a serious and common psychological disorder that requires early diagnosis and treatment. In severe episodes the condition may result in suicidal thoughts. Recently, the need for building an effective audio-based Automatic Depression Detection (ADD) system has sparked the interest of the research community. To date, most of the reported approaches to recognize depression rely on hand-crafted feature extraction for audio data representation. They combine wide variety of audio-related features to improve the classification performance. However, combining many hand-crafted features including relevant and less-relevant can enlarge the feature space which can lead to high-dimensionality issues as not all the features would carry significant information regarding depression. Having high number of features can make the pattern recognition more difficult and increase the risk of overfitting. To overcome these limitations, an audio-based framework of depression detection which includes an adaptation of a deep learning (DL) technique is proposed to automatically extract the highly relevant and compact feature set. This proposedHighlights: A novel audio-based depression detection system using Convolutional Autoencoder. Convolutional Autoencoder for extracting highly correlated and compact feature set. Thorough experimental study based on a real-world depression detection dataset. Complete comparison of proposed feature extraction method with other techniques. Abstract: Depression is a serious and common psychological disorder that requires early diagnosis and treatment. In severe episodes the condition may result in suicidal thoughts. Recently, the need for building an effective audio-based Automatic Depression Detection (ADD) system has sparked the interest of the research community. To date, most of the reported approaches to recognize depression rely on hand-crafted feature extraction for audio data representation. They combine wide variety of audio-related features to improve the classification performance. However, combining many hand-crafted features including relevant and less-relevant can enlarge the feature space which can lead to high-dimensionality issues as not all the features would carry significant information regarding depression. Having high number of features can make the pattern recognition more difficult and increase the risk of overfitting. To overcome these limitations, an audio-based framework of depression detection which includes an adaptation of a deep learning (DL) technique is proposed to automatically extract the highly relevant and compact feature set. This proposed framework uses an end-to-end Convolutional Neural Network-based Autoencoder (CNN AE) technique to learn the highly relevant and discriminative features from raw sequential audio data, and hence to detect depressed people more accurately. In addition, to address the sample imbalance problem we use a cluster-based sampling technique which highly reduces the risk of bias towards the major class (non-depressed). To evaluate the performance and effectiveness of the proposed pipeline, we perform the experiments on Distress Analysis Interview Corpus-Wizard of Oz (DAIC-WOZ) dataset and compare them with the hand-crafted feature extraction methods and other outstanding studies in this domain. The results show that proposed method outperforms other well-known audio-based ADD models with at least 7% improvement in F-measure for classifying depression. … (more)
- Is Part Of:
- Expert systems with applications. Volume 189(2022)
- Journal:
- Expert systems with applications
- Issue:
- Volume 189(2022)
- Issue Display:
- Volume 189, Issue 2022 (2022)
- Year:
- 2022
- Volume:
- 189
- Issue:
- 2022
- Issue Sort Value:
- 2022-0189-2022-0000
- Page Start:
- Page End:
- Publication Date:
- 2022-03-01
- Subjects:
- Audio depression detection -- Semi-supervised learning -- Convolutional Autoencoder -- Early depression detection
Expert systems (Computer science) -- Periodicals
Systèmes experts (Informatique) -- Périodiques
Electronic journals
006.33 - Journal URLs:
- http://www.sciencedirect.com/science/journal/09574174 ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.eswa.2021.116076 ↗
- Languages:
- English
- ISSNs:
- 0957-4174
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 3842.004220
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 20028.xml