These models were created to advance automatic phonetic transcription (APT) beyond the training transcription accuracy.
The workflow to improve APT is called Selective Augmentation and was developed by Tobias Bystrich at Fraunhofer Institute IAIS and using resources of WestAI:
Simulations were performed with computing resources granted by WestAI under project rwth1594.
The models in this project are the reference (RM), helper (HM), baseline (BM) and target model (TM) for the selective augmentation workflow. Additionally, for reimplementation, the provided list of training segments ensures that the RM can predict the highest quality reference transcriptions.
The RM closely corresponds to a reimplemented MultIPA model (
https://github.com/ctaguchi/multipa).
The target model has greatly improved plosive phonation information when measured against the baseline model. This is achieved by augmenting the baseline training data with reliable phonation information from a Hindi helper model.
For more in-depth discussions about this approach and automatic phonetic transcriptions in general, you may consult Tobias Bystrich's master's thesis:
"Multilingual Automatic Phonetic Transcription – a Linguistic Investigation of its Performance on German and Approaches to Improving the State of the Art".
https://doi.org/10.24406/publica-4418