Downloads · 30 days
26
25% of all-time downloads
aitytech/Whisper-Small-MLX-FP16
Whisper-Small-MLX-FP16 is a automatic speech recognition model from aitytech. Use it when you need speech turned into text. It is set up for mlx. The card lists the license as apache-2.0.
This is the OpenAI Whisper Small model converted to MLX format with FP16 precision, optimized for Apple Silicon inference.
Downloads · 30 days
26
25% of all-time downloads
All-time downloads
106
Public
Parameters
241M
481 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors481 MB · 100%
From the Hugging Face model README
This is the OpenAI Whisper Small model converted to MLX format with FP16 precision, optimized for Apple Silicon inference.
| Property | Value |
|---|---|
| Base Model | openai/whisper-small |
| Parameters | ~244M |
| Format | MLX SafeTensors (FP16) |
| Model Size | 458.92 MB |
| Sample Rate | 16,000 Hz |
| Audio Layers | 12 |
| Text Layers | 12 |
| Hidden Size | 768 |
| Attention Heads | 12 |
| Vocabulary Size | 51,865 |
This model is optimized for on-device automatic speech recognition (ASR) on Apple Silicon devices (Mac, iPhone, iPad). It is designed for use with the WhisperKit or MLX frameworks.
config.json - Model configurationmodel.safetensors - Model weights in SafeTensors format (FP16)multilingual.tiktoken - Tokenizerimport mlx_whisper
result = mlx_whisper.transcribe(
"audio.mp3",
path_or_hf_repo="aitytech/Whisper-Small-MLX-FP16",
)
print(result["text"])