Num. Examples
30,000
Num. Channels
1
Num. Classes
10
AudioMNIST consists of audio recordings of 60 different speakers saying the digits 0 to 9, with 50 recordings per digit per speaker [1, 2]. The speakers are a mixture of ages and genders. The recordings are single channel have a sampling rate of 48 kHz. The learning… See the full description on the dataset page:
https://huggingface.co/datasets/monster-monash/AudioMNIST-DS.