Views
No views yet
| Property | Value |
|---|---|
| Base Model | openai/whisper-medium |
| Parameters | ~769M |
| Format | MLX SafeTensors (FP16) |
| Model Size | 1,454.10 MB |
| Sample Rate | 16,000 Hz |
| Audio Layers | 24 |
| Text Layers | 24 |
| Hidden Size | 1024 |
| Attention Heads | 16 |
| Vocabulary Size | 51,865 |
config.json - Model configurationmodel.safetensors - Model weights in SafeTensors format (FP16)multilingual.tiktoken - Tokenizer1import mlx_whisper
2
3result = mlx_whisper.transcribe(
4 "audio.mp3",
5 path_or_hf_repo="aitytech/Whisper-Medium-MLX-FP16",
6)
7print(result["text"])