This is a merged speech dataset containing 11930 audio segments from 24 source datasets.
Total Segments: 11930
Speakers: 69
Languages: tr
Emotions: angry, happy, neutral, sad
Original Datasets: 24
audio: Audio file (WAV format, 16kHz sampling rate)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected… See the full description on the dataset page:
https://huggingface.co/datasets/Codyfederer/testtrdataset-1.