Downloads · 30 days
162
25% of all-time downloads
speech-seq2seq/wav2vec2-2-roberta-large
wav2vec2-2-roberta-large is a automatic speech recognition model from speech-seq2seq. Use it when you need speech turned into text. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
162
25% of all-time downloads
All-time downloads
644
Public
Repo size
98.1 GB
Likes
0
Public
Click a slice to open those files.
.bin3.2 GB · 100%
From the Hugging Face model README
This model was trained from scratch on the librispeech_asr dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer |
|---|---|---|---|---|
| 6.5774 | 0.28 | 500 | 10.5449 | 1.0 |
| 6.706 | 0.56 | 1000 | 9.4411 | 1.0 |
| 6.9182 | 0.84 | 1500 | 10.9554 | 1.0 |
| 6.7416 | 1.12 | 2000 | 10.0801 | 1.0 |
| 6.8778 | 1.4 | 2500 | 9.8569 | 1.0 |
| 6.7694 | 1.68 | 3000 | 10.4234 | 1.0 |
| 6.7415 | 1.96 | 3500 | 10.6545 | 1.0 |
| 6.5997 | 2.24 | 4000 | 10.4268 | 1.0 |
| 6.7672 | 2.52 | 4500 | 11.1929 | 1.0 |
| 6.5254 | 2.8 | 5000 | 12.2365 | 1.0 |