Downloads · 30 days
8
1% of all-time downloads
MatsRooth/wav2vec2-base_down_on
wav2vec2-base_down_on is a audio classification model from MatsRooth. Use it for the audio classification task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
8
1% of all-time downloads
All-time downloads
535
Public
Repo size
757 MB
Likes
0
Public
Click a slice to open those files.
.bin378 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of facebook/wav2vec2-base on the MatsRooth/down_on dataset. It achieves the following results on the evaluation set:
Binary classifier using facebook/wav2vec2/base for the words "down" and "on".
This is a demo of binary audio classification that illustrates data layout, training and evaluation using python and slurm.
The data are utterances of "down" and "on" in superb ks. See down_on_copy.py for the subsetting. This puts wav files in locations
like down_on/data/train/on and down_on/data/train/down
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Accuracy |
|---|---|---|---|---|
| 0.6089 | 1.0 | 29 | 0.1385 | 0.9962 |
| 0.1289 | 2.0 | 58 | 0.0510 | 0.9962 |
| 0.0835 | 3.0 | 87 | 0.0433 | 0.9885 |
| 0.0605 | 4.0 | 116 | 0.0330 | 0.9923 |
| 0.0479 | 5.0 | 145 | 0.0273 | 0.9904 |