This dataset is a relabeled version of the "AI-trainingset voor Named Entity Recognition (NER)" created during the crowdsourcing project "Tag de Tekst" on VeleHanden.nl in 2020. It has been adapted for use in Named Entity Recognition (NER) tasks, with relabeling conducted using Google Deepmind's Gemini 2.0 Flash model.
Original Source: Transcriptions of Dutch notarial texts from the 17th to 19th centuries.… See the full description on the dataset page:
https://huggingface.co/datasets/TimKoornstra/dutch-notarial-ner.