A verse-aligned parallel corpus of Quran translations into languages of the
Caucasus, plus the Arabic original, three Russian translations, and an English
translation as pivot/anchor texts. Built for NMT research on low-resource Caucasian languages: several of
these languages (Lezgian, Kabardian, Adyghe, Dargwa, Karachay-Balkar) have very
little other parallel data of this size and quality.
All texts follow the canonical… See the full description on the dataset page:
https://huggingface.co/datasets/AlidarAsvarov/quran-caucasus.