The dataset is translated from Standford Alpaca instruction dataset via Google Translations API.
Manually fixed the translation error.
Common names and places of Kazakhstan were added.
Intructions of kazakhstan history and cultures were added.
This dataset is curated to fine-tune the LLaMA 2 model for the Kazakh language. It aims to enhance the model's… See the full description on the dataset page:
https://huggingface.co/datasets/AmanMussa/kazakh-instruction-v2.