This is the version 3 of the Galgame-Llasa-3B, a Text-to-Speech (TTS) model fine-tuned for Japanese. This model is based on
HKUSTAudio/Llasa-3B.
This update leads to more consistent and accurate speech synthesis, further improving upon the advances made in v2.
Version 2 was trained on a larger and more diverse dataset, including the original Galgame dataset and other sources.