Downloads · 30 days
21
8% of all-time downloads
aitytech/Whisper-Base-MLX-FP16
Whisper-Base-MLX-FP16 is a automatic speech recognition model from aitytech. Use it when you need speech turned into text. It is set up for mlx. The card lists the license as apache-2.0.
This is the OpenAI Whisper Base model converted to MLX format with FP16 precision, optimized for Apple Silicon inference.
Downloads · 30 days
21
8% of all-time downloads
All-time downloads
261
Public
Parameters
71.8M
144 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors144 MB · 99%
From the Hugging Face model README
This is the OpenAI Whisper Base model converted to MLX format with FP16 precision, optimized for Apple Silicon inference.
| Property | Value |
|---|---|
| Base Model | openai/whisper-base |
| Parameters | ~74M |
| Format | MLX SafeTensors (FP16) |
| Model Size | 137.02 MB |
| Sample Rate | 16,000 Hz |
| Audio Layers | 6 |
| Text Layers | 6 |
| Hidden Size | 512 |
| Attention Heads | 8 |
| Vocabulary Size | 51,865 |
This model is optimized for on-device automatic speech recognition (ASR) on Apple Silicon devices (Mac, iPhone, iPad). It is designed for use with the WhisperKit or MLX frameworks.
config.json - Model configurationmodel.safetensors - Model weights in SafeTensors format (FP16)multilingual.tiktoken - Tokenizerimport mlx_whisper
result = mlx_whisper.transcribe(
"audio.mp3",
path_or_hf_repo="aitytech/Whisper-Base-MLX-FP16",
)
print(result["text"])