Downloads · 30 days
6
14% of all-time downloads
jun-han/Whisper-VAD-Small-Deep-Sparse-squeezeformer
Whisper-VAD-Small-Deep-Sparse-squeezeformer is a automatic speech recognition model from jun-han. Use it when you need speech turned into text. It is set up for transformers. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
6
14% of all-time downloads
All-time downloads
43
Public
Parameters
323M
23.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.3 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of openai/whisper-small on the Voice_Data_Collection_second_edition dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Cer |
|---|---|---|---|---|
| 2.999 | 0.7709 | 2500 | 2.9936 | 101.4092 |
| 2.7197 | 1.5418 | 5000 | 2.7733 | 97.9264 |
| 0.9736 | 2.3127 | 7500 | 1.0609 | 62.7590 |
| 0.5813 | 3.0836 | 10000 | 0.7054 | 42.3849 |
| 0.4856 | 3.8545 | 12500 | 0.5940 | 35.6714 |
| 0.3203 | 4.6253 | 15000 | 0.5318 | 32.1755 |
| 0.2468 | 5.3962 | 17500 | 0.5106 | 30.4677 |
| 0.1902 | 6.1671 | 20000 | 0.4928 | 29.4300 |