This is ContextDep (Context-Dependent Paraphrases in news interviews), created as described in
https://arxiv.org/abs/2404.06670.
This dataset consists of 601 interview turns between a guest and a host of a news interview and 5581 crowd-sourced annotations of whether the host paraphrased the guest. If the annotator classifies a paraphrase, they also provided spans of words where the paraphrase occurs. There are between 3-21 annotations per item. Generally, an item has more annotations if… See the full description on the dataset page:
https://huggingface.co/datasets/AnnaWegmann/Paraphrases-in-Interviews.