Downloads · 30 days
12
4% of all-time downloads
mlx-community/VibeVoice-ASR-5bit
VibeVoice-ASR-5bit is a automatic speech recognition model from mlx-community. Use it when you need speech turned into text. It is set up for mlx-audio. The card lists the license as mit.
This model was converted to MLX format from microsoft/VibeVoice-ASR using mlx-audio version 0.3.0.
Downloads · 30 days
12
4% of all-time downloads
All-time downloads
329
Public
Parameters
8.3B
6.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors6.7 GB · 100%
How the weights are stored.
U327.6B · 91%
From the Hugging Face model README
This model was converted to MLX format from microsoft/VibeVoice-ASR using mlx-audio version 0.3.0.
Refer to the original model card for more details on the model.
pip install -U mlx-audio
python -m mlx_audio.stt.generate --model mlx-community/VibeVoice-ASR-5bit --audio "audio.wav"
from mlx_audio.stt.utils import load_model
from mlx_audio.stt.generate import generate_transcription
model = load_model("mlx-community/VibeVoice-ASR-5bit")
transcription = generate_transcription(
model=model,
audio_path="path_to_audio.wav",
output_path="path_to_output.txt",
format="txt",
verbose=True,
)
print(transcription.text)