Synthetic Multilingual PII NER Dataset
Models Trained Using this Dataset
E3-JSI/gliner-multi-pii-domains-v1
Description
This is a synthetic dataset created for the purposes for training multilingual personally identifiable information (PII) named entity recognition (NER) models.
The examples were generated using a prompt that generates the text and the entities present in the text. In addition, the generated response had to follow the restrictions: