Downloads · 30 days
0
techiaith/whisper-base-ft-commonvoice-cy-en-cpp
whisper-base-ft-commonvoice-cy-en-cpp is a automatic speech recognition model from techiaith. Use it when you need speech turned into text. The card lists the license as apache-2.0.
This model is a version of the openai/whisper-base model, fine-tuned on the techiaith/commonvoice180cyen dataset, and then converted for use in whisper.cpp. Whispercpp is
Downloads · 30 days
0
Access
Public
Updated Nov 8, 2024
Repo size
148 MB
Likes
0
Public
Click a slice to open those files.
.bin148 MB · 100%
From the Hugging Face model README
This model is a version of the openai/whisper-base model, fine-tuned on the techiaith/commonvoice_18_0_cy_en dataset, and then converted for use in whisper.cpp. Whispercpp is a C/C++ port of Whisper that provides high performance inference on offline hardware such as desktops, laptops and mobile devices.
The model is a smaller in size to the corresponding cloud hosted model techiaith/whisper-large-v3-ft-cv-cy-en. It achieves a success rate of 98.34% on detecting the correct language in speech, while for transcribing it achieves the following WER results:
whispercpp makes it easy to use models in many platforms and applications. See the 'examples' folder in the whispercpp github repo for more information and example code.
To get quickly started with whispercpp's basic usage however, follow the 'Quick Start' but download this model with the following command:
$ wget https://huggingface.co/techiaith/whisper-base-ft-cv-cy-en-cpp/resolve/main/ggml-model.bin