Postprocessing
Training
Published datasets (union, merged union) and models (EVENT, LOC, MISC, ORG, PER, TIME)
A named entity recognition system (NER) was trained on text extracted from Oberdeutsche Allgemeine Litteraturueitung (OALZ) of the first quarter (January, Febuary, March) of 1788. The scans from which text was extracted can be found at Bayerische Staatsbibliothek using the extraction strategy of the KEDiff project, which can be found at cborgelt/KEDiff.… See the full description on the dataset page:
https://huggingface.co/datasets/LelViLamp/oalz-1788-q1-ner-annotations-union-dataset.