Dusha is a bi-modal corpus suitable for speech emotion recognition (SER) tasks.
The dataset consists of audio recordings with Russian speech and their emotional labels.
The corpus contains approximately 350 hours of data. Four basic emotions that usually appear in a dialog with
a virtual assistant were selected: Happiness (Positive), Sadness, Anger and Neutral emotion.