This model is a fine-tuned version of
gpt2 on the
wikitext.
It achieves the following results on the evaluation set:
This is a practical hands-on experience for better understanding 🤗 Transformers and 🤗 Datasets. This model is GPT2(124M) fine-tuned on wikitext(103-raw-v1) on 1 x RTX 4090.