Dataset Card for Fragment Of Bookcorpus
Dataset Description
A smaller sample of the bookcorpus dataset, Which includes around 100,000 lines of text.
^^^(In comparison to the original bookcorpus' 74.1~ Million lines of text)^^^
Dataset Summary
Modified and Uploaded to the hugggingface library as a part of a project. Essentially aiming at Open-Ended conversation data.
This dataset is basically a fragment of the infamous bookcorpus dataset.
Which aims to… See the full description on the dataset page: https://huggingface.co/datasets/Seraphiive/FragmentOfBOOKCORPUS.