This dataset is a curated collection of sentence-length paraphrases derived from two primary sources:
humarin/chatgpt-paraphrases
xwjzds/paraphrase_collections.
Dataset Details
Dataset Description
The dataset is structured to provide pairs of sentences from an original text and its paraphrase(s). For each entry:
The "text" field contains the least readable paraphrase.
The "paraphrase" field contains the most readable paraphrase.… See the full description on the dataset page:
https://huggingface.co/datasets/agentlans/sentence-paraphrases.