Downloads · 30 days
6
1% of all-time downloads
luigisaetta/whisper-medium-it
whisper-medium-it is a automatic speech recognition model from luigisaetta. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
This model is a fine-tuned version of openai/whisper-medium on the commonvoice110 dataset.
Downloads · 30 days
6
1% of all-time downloads
All-time downloads
717
Public
Repo size
18.3 GB
Likes
2
Public
Click a slice to open those files.
.bin3.1 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of openai/whisper-medium on the common_voice_11_0 dataset.
It achieves the following results on the evaluation set:
This model is a fine-tuning of the OpenAI Whisper Medium model, on the specified dataset.
This model has been developed as part of the Hugging Face Whisper Fine Tuning sprint, December 2022.
It is meant to spread the knowledge on how these models are built and can be used to develop solutions where it is needed ASR on the Italian Language.
It has not been extensively tested. It is possible that on other datasets the accuracy will be lower.
Please, test it before using it.
Trained and tested on Mozilla Common Voice, vers. 11
The script run.sh, and the Python file, used for the training are saved in the repository.
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 0.1216 | 0.2 | 1000 | 0.2289 | 10.0594 |
| 0.1801 | 0.4 | 2000 | 0.1851 | 7.6593 |
| 0.1763 | 0.6 | 3000 | 0.1615 | 6.5258 |
| 0.1337 | 0.8 | 4000 | 0.1506 | 6.0427 |
| 0.0742 | 1.05 | 5000 | 0.1452 | 5.7191 |