This is a merged speech dataset containing 345 audio segments from 2 source datasets.
Total Segments: 345
Speakers: 7
Languages: en
Emotions: neutral, angry, happy, sad
Original Datasets: 2
audio: Audio file (WAV format, 16kHz sampling rate)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected emotion… See the full description on the dataset page:
https://huggingface.co/datasets/Codyfederer/last.