Views
No views yet
mlx-community/Qwen3-TTS-12Hz-1.7B-Base-8bit
at revision e7dd0585652209fa0d7783659aad4e8a324de11c (itself an MLX conversion of the
corresponding Qwen/Qwen3-TTS checkpoint), with one change:
the 622 MB BF16 talker.model.text_embedding tensor is quantized to affine 8-bit
(group size 64) by Vocello's pinned conversion tooling
(scripts/convert_text_embedding_8bit.py, python-mlx 0.32.0). Every other tensor is
byte-identical to the source revision.benchmarks/OPTIMIZATION.md §N).