Downloads · 30 days
8
1% of all-time downloads
chrisjay/afrospeech-wav2vec-run
afrospeech-wav2vec-run is a audio classification model from chrisjay. Use it for the audio classification task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
This model is a fine-tuned version of facebook/wav2vec2-base on the crowd-speech-africa, which was a crowd-sourced dataset collected using the afro-speech Space.
Downloads · 30 days
8
1% of all-time downloads
All-time downloads
638
Public
Repo size
757 MB
Likes
0
Public
Click a slice to open those files.
.bin378 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-base on the crowd-speech-africa, which was a crowd-sourced dataset collected using the afro-speech Space.
The model was trained on a mixed audio data from Rundi (run).
Below is a distribution of the dataset (training and valdation)

It achieves the following results on the validation set:
The confusion matrix below helps to give a better look at the model's performance across the digits. Through it, we can see the precision and recall of the model as well as other important insights.

The following hyperparameters were used during training:
| Training Loss | Epoch | Validation Accuracy |
|---|---|---|
| 0.00183 | 1 | 0.6 |
| 0.0003991 | 50 | 0.8 |
| 0.0002174 | 100 | 0.6 |
| 0.0043911 | 150 | 0.4 |