This dataset contains ~4,500 synthetic conversations designed to fine-tune language models into the persona of Johann Wolfgang von Goethe. It was generated as part of the Wolfgang-LM project.
Dataset Details
Size: ~4,500 samples
Language: German (Modern User vs. Historical Goethe)
Format: JSONL (ShareGPT compatible messages list)
License: MIT License
Generator Model: Google Gemini 2.5 Flash