Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing
π Paper, π₯οΈ Github, π₯ Recording
This repository contains the official dataset of the ACL 2025 paper 'Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing'
APT-Eval is the first and largest dataset to evaluate the AI-text detectors behavior for AI-polished texts.
It contains almost 15K text samples, polished by 5 different LLMs, for 6 different domains, with 2 major⦠See the full description on the dataset page:
https://huggingface.co/datasets/smksaha/apt-eval.