This dataset, german-level-tuner, is designed for fine-tuning language models to assess the German language proficiency of a given text, based on the Common European Framework of Reference for Languages (CEFR).The dataset consists of German texts labeled with their corresponding CEFR level (A1, A2, B1, B2, C1).
The primary goal of this dataset is to enable the development of models that can automatically classify the difficulty of German texts, making it a… See the full description on the dataset page:
https://huggingface.co/datasets/AlbertoB12/german-level-tuner.