Downloads · 30 days
9
14% of all-time downloads
scasutt/wav2vec2-base_toy_train_data_augment_0.1
wav2vec2-base_toy_train_data_augment_0.1 is a automatic speech recognition model from scasutt. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
9
14% of all-time downloads
All-time downloads
66
Public
Repo size
4.2 GB
Likes
0
Public
Click a slice to open those files.
.bin378 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-base on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 3.1342 | 1.05 | 250 | 3.3901 | 0.9954 |
| 3.0878 | 2.1 | 500 | 3.4886 | 0.9954 |
| 3.0755 | 3.15 | 750 | 3.4616 | 0.9954 |
| 3.0891 | 4.2 | 1000 | 3.5316 | 0.9954 |
| 3.0724 | 5.25 | 1250 | 3.2608 | 0.9954 |
| 3.0443 | 6.3 | 1500 | 3.3881 | 0.9954 |
| 3.0421 | 7.35 | 1750 | 3.4507 | 0.9954 |
| 3.0448 | 8.4 | 2000 | 3.4525 | 0.9954 |
| 3.0455 | 9.45 | 2250 | 3.3342 | 0.9954 |
| 3.0425 | 10.5 | 2500 | 3.3385 | 0.9954 |
| 3.0457 | 11.55 | 2750 | 3.4411 | 0.9954 |
| 3.0375 | 12.6 | 3000 | 3.4459 | 0.9954 |
| 3.0459 | 13.65 | 3250 | 3.3883 | 0.9954 |
| 3.0455 | 14.7 | 3500 | 3.3417 | 0.9954 |
| 3.0524 | 15.75 | 3750 | 3.3908 | 0.9954 |
| 3.0443 | 16.81 | 4000 | 3.3932 | 0.9954 |
| 3.0446 | 17.86 | 4250 | 3.4052 | 0.9954 |
| 3.0412 | 18.91 | 4500 | 3.3776 | 0.9954 |
| 3.0358 | 19.96 | 4750 | 3.3786 | 0.9954 |