Views
No views yet
onnxruntime-node with DirectML/CUDA).torchaudio.pipelines.MMS_FA), part of the
Massively Multilingual Speech (MMS) project.Wav2Vec2ForCTC to ONNX (opset 17, dynamic batch/samples axes,
input: input_values [batch, samples] float32, output: logits [batch, frames, 31] float32).keep_io_types=True, so I/O stays float32).| file | size | sha256 |
|---|---|---|
mms_300m_fa_fp16.onnx | 631,499,926 bytes | 9823aa1e22dfc91fe1b1185ead5db5cab4512ce604ed18459b17a9eb92a00cd2 |
vocab.json | — | token vocabulary + blank id + frame metadata |
<blank>). Vocabulary is romanized (lowercase a–z + apostrophe).