-
text-generation
language:
-
rw
size_categories:
-
1K<n<10K
The Kinyarwanda Monolingual Dataset version 1 is a large collection of Kinyarwanda language texts aimed at supporting the development of NLP and AI applications which can process Kinyarwanda texts.
This dataset contains 1,068,161, with 63,001,765 words and includes diverse content types such as news articles, government reports, religious texts, legal documents, educational… See the full description on the dataset page:
https://huggingface.co/datasets/mbazaNLP/kinyarwanda_monolingual_v01.1.