Project Gutenberg Temporal Corpus
Repository Updates
02.09.2025
Fix the unsafe issue in the retrieved contents files.
Add the detailed Generes-Super_Generes Mapping in metadata files.
Usage
To use this dataset, we suggest cloning the repository and accessing the files directly. The dataset is organized into several zip files and CSV files, which can be easily extracted and read using standard data processing libraries in Python or other programming… See the full description on the dataset page: https://huggingface.co/datasets/Texttechnologylab/project-gutenberg-temporal-corpus.