This corpus contains paired data of speech, articulatory movements and phonemes. There are 38 speakers in the corpus, each with 460 utterances.
The raw audio files are in audios.zip. The ema data and preprocessed data is stored in processed.zip. The processed data can be loaded with pytorch and has the following keys -
ema_trimmed_and_normalised_with_6_articulators: The ema data after trimming… See the full description on the dataset page:
https://huggingface.co/datasets/sshahnawazuddin/SPIRE_EMA_CORPUS.