Downloads · 30 days
56
3% of all-time downloads
bardsai/whisper-medium-pl
whisper-medium-pl is a automatic speech recognition model from bardsai. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
56
3% of all-time downloads
All-time downloads
2.1K
Public
Parameters
764M
19.9 GB on disk
Likes
2
Public
Click a slice to open those files.
.bin3.1 GB · 50%
From the Hugging Face model README
This model is a fine-tuned version of openai/whisper-medium on the Common Voice 11.0 dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 0.0474 | 1.1 | 1000 | 0.2561 | 9.4612 |
| 0.0119 | 3.09 | 2000 | 0.2901 | 8.9726 |
| 0.0045 | 5.08 | 3000 | 0.3151 | 8.8870 |
| 0.0007 | 7.07 | 4000 | 0.4218 | 8.6032 |
| 0.0005 | 9.07 | 5000 | 0.3739 | 8.5898 |
When tested on diffrent polish ASR datasets (splits: test), this model achieves the following results:
| Dataset | WER | WER unnormalized | CER | MER |
|---|---|---|---|---|
| common_voice_11_0 | 8.85 | 21.75 | 2.63 | 8.76 |