This dataset contains 707,261 parallel examples specifically curated for training Grammatical Error Correction (GEC) models for the Russian language. The dataset follows an instruction-tuning format, making it suitable for fine-tuning instruction-following language models.
input
string
Source text containing grammatical… See the full description on the dataset page:
https://huggingface.co/datasets/p1746-lingua/ru-gec-v1.