This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at
https://creativecommons.org/licenses/by-nc/4.0/.
A large-scale Twi language dataset containing clean and naturally sounding Twi text across
four distinct styles — monologue, narrative, dialogue, and storytelling — generated
from real Ghanaian news topics as… See the full description on the dataset page:
https://huggingface.co/datasets/ghananlpcommunity/pristine-twi.