Downloads · 30 days
136
27% of all-time downloads
aystream/GigaAM-v3-e2e-ctc-mlx
GigaAM-v3-e2e-ctc-mlx is a automatic speech recognition model from aystream. Use it when you need speech turned into text. It is set up for mlx. The card lists the license as mit.
MLX port of GigaAM-v3 for fast Russian speech recognition on Apple Silicon. 180x realtime on M2 Max.
Downloads · 30 days
136
27% of all-time downloads
All-time downloads
496
Public
Repo size
884 MB
Likes
0
Public
Click a slice to open those files.
.safetensors884 MB · 100%
From the Hugging Face model README
MLX port of GigaAM-v3 for fast Russian speech recognition on Apple Silicon. 180x realtime on M2 Max.
pip install gigaam-mlx
from gigaam_mlx import load_model, transcribe
model, tokenizer = load_model() # downloads weights automatically
text = transcribe(model, tokenizer, "recording.wav")
print(text)
Or via CLI:
gigaam-mlx recording.wav
MacBook Pro M2 Max, 20-second chunk:
| Backend | Time | Realtime |
|---|---|---|
| MLX CTC (this) | 0.11s | 180x |
| PyTorch MPS RNNT | 0.76s | 26x |
| ONNX CPU CTC | 1.66s | 12x |