Downloads · 30 days
8
12% of all-time downloads
scasutt/wav2vec2-base_toy_train_data_random_high_pass
wav2vec2-base_toy_train_data_random_high_pass is a automatic speech recognition model from scasutt. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
8
12% of all-time downloads
All-time downloads
68
Public
Repo size
4.2 GB
Likes
0
Public
Click a slice to open those files.
.bin378 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-base on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 3.061 | 2.1 | 500 | 3.0551 | 1.0 |
| 1.1294 | 4.2 | 1000 | 1.3102 | 0.8777 |
| 0.7051 | 6.3 | 1500 | 1.2081 | 0.8092 |
| 0.5421 | 8.4 | 2000 | 1.2280 | 0.7684 |
| 0.448 | 10.5 | 2500 | 1.2459 | 0.7506 |
| 0.3777 | 12.6 | 3000 | 1.3533 | 0.7631 |
| 0.3611 | 14.7 | 3500 | 1.2058 | 0.7291 |
| 0.3177 | 16.81 | 4000 | 1.3168 | 0.7185 |
| 0.279 | 18.91 | 4500 | 1.2841 | 0.7222 |