Koumankan4Dyula: Parallel Dyula - French Dataset for Machine Learning
Overview
The Koumankan4Dyula corpus consists of 10929 pairs of Dioula-French sentences.
This corpus is part of the Koumankan project, which proposes a scalable and cost-effective method for extending the CommonVoice dataset to the Dyula language and other African languages.