AI vs Human dataset on the CNN DailyNews
Dataset Description
This dataset showcases pairs of truncated text and their respective completions, crafted either by humans or an AI language model.
Each article was randomly truncated between 25% and 50% of its length.
The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation.