Views
No views yet
Aratako/Irodori-TTS-500M-v3, source revision 236c1e56591279fc24e3c1bf6609fc06e48dde28Aratako/Irodori-TTS-600M-v3-VoiceDesign, source revision e863a3a93e652e09afeff3e84823a206a0a60314Aratako/Semantic-DACVAE-Japanese-32dimllm-jp/llm-jp-3-150m500m-v3/: standard Japanese TTS and reference-audio voice cloning.600m-v3-vd/: VoiceDesign model with caption conditioning.context_kv.onnx + dit_step.onnx execution path used by the Onsei iOS app, DACVAE encode/decode, speaker and text encoders, duration prediction, configuration, and tokenizer data. manifest.json records file sizes and SHA-256 digests used by the app to verify downloads.LICENSES/ and THIRD_PARTY_NOTICES.md. Each component remains subject to its own license; no relicensing is implied.1@misc{irodori-tts-v3,
2 author = {Chihiro Arata},
3 title = {Irodori-TTS: A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control},
4 year = {2026},
5 publisher = {Hugging Face},
6 howpublished = {https://huggingface.co/Aratako/Irodori-TTS-500M-v3}
7}