A unified parallel-translation dataset assembled from the sources catalogued in
Datasets_chagatai_final(1) (1).xlsx. Every underlying corpus is normalised to a
single flat schema so it can be loaded and mixed in one line.
src_lang
string
source language code (normalised, see… See the full description on the dataset page:
https://huggingface.co/datasets/Inomjonov/chagatai-parallel.