The toxic-onto-bg dataset consists of 299 Bulgarian manually annotated words from the Flores Toxicity 200 dataset across four categories: toxic language, medical terminology, non-toxic language, and terms related to minority communities,
as well as class formalisms and definitions.
The ontology is aimed for language and media researches and developers of toxic language… See the full description on the dataset page:
https://huggingface.co/datasets/sofia-uni/toxic-onto-bg.