Key result
Deep learning model detects stress from voice recordings with ~97% accuracy.
Why the study?
Stress negatively impacts physical and mental health, and artificial intelligence approaches are increasingly used to detect stress through speech.
A CNN-based deep learning model using Mel Spectrogram and MFCC features can accurately detect stress from voice recordings with 97.1% accuracy.
May support non-invasive stress screening; leaves open prospective validation before clinical adoption.
Stress, a change in psychological reactions from a calm state to an emotional state, is a psychological problem that can negatively impact a person's physical and mental condition. Daily life that is full of pressures can be a stressor that triggers the stress. Various artificial intelligence-based technological approaches are currently used to detect stress through various indicators, one of which is using speech. In this study, a deep learning model based on CNN architecture was developed to detect stress through voice recording using various sound features extracted in the signal domain. The performance evaluation of the model was demonstrated using an open-source dataset (Crema-D and TESS), and the best accuracy value obtained was 97.1% in performing binary classification on stressed and unstressed labelled speech. The highest accuracy was obtained from experiments using various combinations of sound features in the signal domain using a combination of Mel Spectrogram and MFCC features. This evaluation result shows that the deep learning model with the appropriate sound feature extraction can accurately detect stress through voice recording.
No takes yet. Share an insight, caveat, or question.
Chyan et al. (2022) studied Stress. Deep learning model based on CNN architecture using Mel Spectrogram and MFCC features was evaluated on Binary classification accuracy on stressed and unstressed labelled speech. A deep learning model based on CNN architecture using a combination of Mel Spectrogram and MFCC features achieved 97.1% accuracy in detecting stress through voice recordings.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: