This dataset was created while evaluating and comparing the models trained with Listen Attend Understand regularization and our E2E-ST model.
The audio is from jeli-asr test set; the regularization loss weight lambda in the paper is represented by the character "k" in the fields of this dataset, each field represent a model with a specific decoding strategy (CTC or TDT)