Downloads · 30 days
24
6% of all-time downloads
mlx-community/VibeVoice-ASR-6bit
VibeVoice-ASR-6bit is a automatic speech recognition model from mlx-community. Use it when you need speech turned into text. It is set up for mlx-audio. The card lists the license as mit.
This model was converted to MLX format from microsoft/VibeVoice-ASR using mlx-audio version 0.3.0rc1.
Downloads · 30 days
24
6% of all-time downloads
All-time downloads
417
Public
Parameters
8.3B
7.6 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors7.6 GB · 100%
How the weights are stored.
U327.6B · 91%
From the Hugging Face model README
This model was converted to MLX format from microsoft/VibeVoice-ASR using mlx-audio version 0.3.0rc1.
Refer to the original model card for more details on the model.
pip install -U mlx-audio
python -m mlx_audio.stt.generate --model mlx-community/VibeVoice-ASR-6bit --audio "audio.wav"
from mlx_audio.stt.utils import load_model
from mlx_audio.stt.generate import generate_transcription
model = load_model("mlx-community/VibeVoice-ASR-6bit")
transcription = generate_transcription(
model=model,
audio_path="path_to_audio.wav",
output_path="path_to_output.txt",
format="txt",
verbose=True,
)
print(transcription.text)