This dataset consists of 925 sentences in English paired with a broad topic descriptor for use as example data in product demonstrations or student projects.
This data can be loaded using the following Python code.
from datasets import load_dataset
It can then be clustered using the… See the full description on the dataset page:
https://huggingface.co/datasets/billingsmoore/text-clustering-example-data.