KazParC
Kazakh Parallel Corpus (KazParC) is a parallel corpus designed for machine translation across Kazakh, English, Russian, and Turkish. The first and largest publicly available corpus of its kind, KazParC contains a collection of 372,164 parallel sentences covering different domains and developed with the assistance of human translators.
proverbs and sayings
terminology glossaries
phrasebooks
literary works
periodicals
language learning… See the full description on the dataset page:
https://huggingface.co/datasets/issai/kazparc.