This is supervised fine-tuned model for text summarization based on GPT-2 (large). It has been finetuned on the filtered version of TL;DR train dataset, which can be found and downloaded from here:
https://github.com/openai/summarize-from-feedback.
This model has been trained using the TLR library and SFTTrainer class from Huggingface.
Slurm cluster with 8 x H100 Nvidia GPUs.