AfriMMD is a multilingual dataset created to enhance linguistic diversity in AI,
focusing on African languages. This is a proof-of-concept experiment on the use
of multimodal datasets to represent African languages in AI. The dataset contains
translations of the captions in the widely-used Flickr8k dataset into 20 African
languages. The goal is to address the underrepresentation of African languages
in AI and foster more… See the full description on the dataset page:
https://huggingface.co/datasets/AfriMM/AFRICaption.