Downloads · 30 days
5
1% of all-time downloads
tclong/wav2vec2-base-vios
wav2vec2-base-vios is a automatic speech recognition model from tclong. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
1% of all-time downloads
All-time downloads
564
Public
Repo size
4.7 GB
Likes
0
Public
Click a slice to open those files.
.bin1.3 GB · 58%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-xls-r-300m on the vivos_dataset dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 4.4755 | 1.37 | 500 | 0.7991 | 0.5957 |
| 0.5424 | 2.75 | 1000 | 0.4290 | 0.3653 |
| 0.3586 | 4.12 | 1500 | 0.3809 | 0.2890 |
| 0.2824 | 5.49 | 2000 | 0.3808 | 0.2749 |
| 0.2249 | 6.87 | 2500 | 0.3467 | 0.2389 |
| 0.1745 | 8.24 | 3000 | 0.3688 | 0.2384 |
| 0.1459 | 9.61 | 3500 | 0.3729 | 0.2427 |