GPT-2 model for Text Generation in luxembourgish language, trained on 711 MB of text data, consisting of RTL.lu news articles, comments, parlament speeches, the luxembourgish Wikipedia, Newscrawl, Webcrawl and subtitles. Created via transfer learning with an German base model, feature space mapping from LB on Base feature space and gradual layer freezing.
The training took place on a 32 GB Nvidia Tesla V100
1from transformers import AutoTokenizer, AutoModelForCausalLM
2tokenizer = AutoTokenizer.from_pretrained("laurabernardy/LuxGPT2-basedGER")
3
4model = AutoModelForCausalLM.from_pretrained("laurabernardy/LuxGPT2-basedGER")
See the
GPT2 model card for considerations on limitations and bias. See the
GPT2 documentation for details on GPT2.