SauerkrautTTS-Preview-0.1 is a fine-tuned Text-to-Speech (TTS) model based on the powerful canopylabs/orpheus-3b-0.1-ft.
This preview model introduces four distinct German-speaking voices—Lena, Anna, Max, and Tom-crafted using original audio recordings captured with a Rhode Studio microphone and Mimic Studio, alongside carefully curated synthetic data. The high quality and careful curation of our overall dataset enable the model to produce clear and natural speech outputs, even from this initial release.
To achieve optimal results, we recommend using a lower temperature for clear and stable outputs. Higher temperatures will enhance dynamism and expressiveness but might introduce instability.
Example inference settings:
temperature = 0.5 # Adjust lower for clearer output, higher for creativity
Future Plans
This model represents our first exploratory step into advanced German-language TTS. Expect significant improvements in upcoming versions, including:
Enhanced voice clarity
Expanded speaker diversity
Greater stability across temperature ranges
Stay tuned for future releases and updates!
License
SauerkrautTTS-Preview-0.1 is openly available under the CC BY-NC 4.0 License, encouraging reuse, remixing, and improvements by the community.
Acknowledgments
We thank Unsloth for their invaluable training script, which we utilized in a lightly modified form for training this model.
Also we are thankful for the German Ministry of Education and Research (BMBF) for funding our Project ARGUS, in which we developed SauerkrautTTS.