This dataset was created in an effort to create a machine translation model for English-to-Kinyarwanda translation and vice-versa in a tourism-geared context.
Repository:link to the GitHub repository containing the code for training the model on this data, and the code for the collection of the monolingual data.
Data Format: TSV
Data Source: web scraping, manual annotation
Model: huggingface model link.
25375 49363 21210 Bird watching… See the full description on the dataset page:
https://huggingface.co/datasets/fair-forward/test.