A conservative, human-adjudicated correction layer over the standard English all-words word sense
disambiguation benchmark.
A panel of models flagged 363 contested items in Maru et al. 2022's
ALL_NEW; three professional lexicographers adjudicated those items independently, changing 211 gold
labels and removing 56. The other 4,554 items carry their source labels unchanged — they were
never reviewed, and lexEN makes no claim about them.
The complete review record is… See the full description on the dataset page:
https://huggingface.co/datasets/GliteTech/lexen.