Downloads · 30 days
285
9% of all-time downloads
Akashpb13/Hausa_xlsr
Hausa_xlsr is a automatic speech recognition model from Akashpb13. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
This model is a fine-tuned version of facebook/wav2vec2-xls-r-300m It achieves the following results on the evaluation set (which is 10 percent of train data set merged with invalidated data, reported, other, and dev…
Downloads · 30 days
285
9% of all-time downloads
All-time downloads
3.1K
Public
Repo size
3.8 GB
Likes
5
Public
Click a slice to open those files.
.bin1.3 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-xls-r-300m It achieves the following results on the evaluation set (which is 10 percent of train data set merged with invalidated data, reported, other, and dev datasets):
"facebook/wav2vec2-xls-r-300m" was finetuned.
More information needed
Training data - Common voice Hausa train.tsv, dev.tsv, invalidated.tsv, reported.tsv and other.tsv Only those points were considered where upvotes were greater than downvotes and duplicates were removed after concatenation of all the datasets given in common voice 7.0
For creating the training dataset, all possible datasets were appended and 90-10 split was used.
The following hyperparameters were used during training:
| Step | Training Loss | Validation Loss | Wer |
|---|---|---|---|
| 500 | 5.175900 | 2.750914 | 1.000000 |
| 1000 | 1.028700 | 0.338649 | 0.497999 |
| 1500 | 0.332200 | 0.246896 | 0.402241 |
| 2000 | 0.227300 | 0.239640 | 0.395839 |
| 2500 | 0.175000 | 0.239577 | 0.373966 |
| 3000 | 0.140400 | 0.243272 | 0.356095 |
| 3500 | 0.119200 | 0.263761 | 0.365164 |
| 4000 | 0.099300 | 0.265954 | 0.353428 |
| 4500 | 0.084400 | 0.276367 | 0.349693 |
| 5000 | 0.073700 | 0.282631 | 0.343825 |
| 5500 | 0.068000 | 0.282344 | 0.341158 |
| 6000 | 0.064500 | 0.281591 | 0.342491 |
mozilla-foundation/common_voice_8_0 with split testpython eval.py --model_id Akashpb13/Hausa_xlsr --dataset mozilla-foundation/common_voice_8_0 --config ha --split test