[!CAUTION]
this dataset is old, go to the regular tuxsentience for the latest one
Version 2 of the tuxsentience dataset.
We operate on the GIGO philosophy (Grain in, Grain out). To maximize Grain we have inserted as much grain as possible into this dataset. This is at least 50% more grain per grain!
Data is manually curated by the GrainWare team based on multiple sources, the main ones being things specifically written… See the full description on the dataset page:
https://huggingface.co/datasets/GrainWare/tuxsentience-old.