Pre-extracted Kyutai Mimi tokens for the
Jenny TTS Dataset —
a single female speaker, ~30h, clean studio-quality recordings. Apache-2.0 license.
Column
Type
Notes
id
string
e.g. jenny_0
text
string
mixed-case with punctuation
codes
int16[k=8][n_frames]
Mimi codebook indices @ 12.5 fps
n_frames
int32
k_codebooks
int32
8
No speaker_id column — single speaker dataset.