This dataset contains 100 English NER records prepared for the Polygraf Applied NLP / NER technical task.
The annotations were made messy on purpose. Some spans are wrong, some labels are inconsistent, some records contain traps, and some records are clean. Candidates should load this dataset, review the annotations carefully, and correct them following the Labels and Baseline labeling rules below, plus any extra policy rules they… See the full description on the dataset page:
https://huggingface.co/datasets/polygraf-ai/applied-nlp-ner-candidate-starter-100.