Corpus for Learners of Estonian as a Second Language 2022 with Synthetic Grammatical Errors
Dataset Summary
The "Corpus for Learners of Estonian as a Second Language 2022" (Eesti keele kui teise keele õppekorpus 2022) is a specialized linguistic resource designed to support learners of Estonian as a second language. The corpus is composed of sentences extracted from 34 different Estonian as a Second Language coursebooks, ranging from A1 to C1 levels. We have introduced… See the full description on the dataset page: https://huggingface.co/datasets/paulpall/textbooks-sentences_estonian.