YoruMed: A Yoruba Medical Terminology Dataset
Dataset Description
YoruMed is a structured Yoruba medical terminology dataset, comprising 1,000 English medical terms, their plain-language English definitions, and corresponding Yoruba translations.
Yoruba is spoken by over 50 million people across Nigeria, Benin, and Togo, yet remains severely underrepresented in biomedical NLP. YoruMed addresses this gap by providing a curated, linguistically annotated dataset that… See the full description on the dataset page: https://huggingface.co/datasets/temitopeolagoke/yorumed.