This dataset is created from Speech Emotion Recognition (en) dataset.
This dataset includes the 4 most popular datasets in English: Crema, Ravdess, Savee, and Tess, containing a total of over 12,000 .wav audio files. Each of these four datasets includes 6 to 8 different emotional labels.
It includes the 7 types of emotions contained in speech.
emotions = ['angry', 'disgust', 'fear', 'happy', 'neutral', 'sad', 'surprise']