A deterministic, text-only corpus of 96,000 original English
Markdown stories built for SmolGPT-Fables. Every row is one complete supervised
story example with an exact prompt / completion boundary, a requested scene
count from one to six, and plain-language conditioning fields.
No model, API, browser, or network service was used to create this dataset.
96,000 stories across 96,000 isolated story families
25 genres and all… See the full description on the dataset page:
https://huggingface.co/datasets/neonforestmist/smolgpt-markdown-stories.