This dataset contains a cleaned, processed, and high-quality collection of Burmese Wikipedia articles, curated to serve as a robust foundation for Natural Language Processing (NLP) and Artificial Intelligence development in the Burmese language.
DatarrX (Burmese: ဒေတာအက်စ်) is a non-profit open-source foundation dedicated to building a robust digital foundation for the Burmese language in the AI era. We believe that… See the full description on the dataset page:
https://huggingface.co/datasets/DatarrX/myanmar-Wikipedia.