The training dataset of this model consists of 23 million tweets in Greek, of approximately 5000 users in total, spanning from 2008 to 2018.
The model has been trained to support the work for the paper
Multimodal Hate Speech Detection in Greek Social Media
1from transformers import AutoTokenizer, AutoModel
2
3tokenizer = AutoTokenizer.from_pretrained("Konstantinos/BERTaTweetGR")
4model = AutoModel.from_pretrained("Konstantinos/BERTaTweetGR")