This dataset contains superficially meaningful English text that entirely lacks global coherence and meaning.
While individual sentences in the scrambled and replacement subsets are grammatically valid, when combined, they do not form a cohesive narrative or logical text.
This dataset is designed to train models on adversarial text classification, natural language inference (NLI), and coherence detection.
The dataset is derived… See the full description on the dataset page:
https://huggingface.co/datasets/agentlans/garbled-text.