An open-source dataset designed for information retrieval and natural language processing tasks.
This dataset is the processed version of reuters-21578 dataset.
Reuters-21578 text categorization test collection
Distribution 1.0 (v 1.2)
26 September 1997
David D. Lewis
AT&T Labs - Research
lewis@research.att.com
The dataset was processed as part of our work on the reuters-search-engine project, where it was my primary… See the full description on the dataset page:
https://huggingface.co/datasets/IsmaelMousa/reuters.