The W7.2 v1 classifier (idirectships/abacus-cheat-tell-v1) achieved perfect accuracy (1.0/1.0/1.0)
on a degenerate task. Its negatives were FineWeb-Edu post-1930 text — stylistically obvious compared
to pre_modern TEAs. The classifier learned modernity detection (modern vocab → anachronism), not
knowledge leakage detection.
The W10 AGI verification verdict requires a classifier that can detect a pre-modern… See the full description on the dataset page: https://huggingface.co/datasets/idirectships/abacus-cheat-tell-eval-v2.