GutenQA consists of book passages manually extracted from Project Gutenberg and subsequently segmented using LumberChunker.
Book Name: The title of the book from which the passage is extracted.
Book ID: A unique integer identifier assigned to each book.
Chunk ID: An integer identifier for each chunk of the book. Chunks are listed in the sequence theyโฆ See the full description on the dataset page:
https://huggingface.co/datasets/LumberChunker/GutenQA.