A multilingual dataset of 69,300 labeled sentences (9,900 per language) across 6 sentence type categories and 7 languages. Designed for training and evaluating sentence-type classifiers in multilingual contexts.
Total entries: 69,300 (9,900 × 7 languages)
Languages: English (EN), Spanish (ES), French (FR), German (DE), Italian (IT), Portuguese (PT), Dutch (NL)
Class distribution: 13,200 entries per… See the full description on the dataset page:
https://huggingface.co/datasets/TigreGotico/sentence-types-multilingual.