Downloads · 30 days
20
100% of all-time downloads
itsskofficial/livewhisper-gu-medium
livewhisper-gu-medium is a automatic speech recognition model from itsskofficial. Use it when you need speech turned into text. It is set up for ctranslate2. The card lists the license as apache-2.0.
A CTranslate2 conversion of vasista22/whisper-gujarati-medium for faster-whisper, as downloaded by LiveWhisper, a voice dictation app for Windows.
Downloads · 30 days
20
100% of all-time downloads
All-time downloads
20
Public
Repo size
1.5 GB
Likes
0
Public
Click a slice to open those files.
.bin1.5 GB · 100%
From the Hugging Face model README
A CTranslate2 conversion of vasista22/whisper-gujarati-medium for faster-whisper, as downloaded by LiveWhisper, a voice dictation app for Windows.
All credit for the model goes to its original authors. This repository only
changes the file format (float16 weights, plus tokenizer.json and
preprocessor_config.json that faster-whisper needs). The licence is the
original's: apache-2.0.
Clearly better than large-v3, but still misses about half the words. Whisper-medium fine-tuned by SPRING Lab, IIT Madras; 1.5 GB converted.
| metric | result |
|---|---|
| FLEURS word error | 49.8% (large-v3 67.7%) |
| delivered Gujlish | 48.5% (large-v3 63.3%) |
from faster_whisper import WhisperModel
model = WhisperModel("itsskofficial/livewhisper-gu-medium", device="cuda", compute_type="int8_float16")