Downloads · 30 days
22
19% of all-time downloads
willopcbeta/whisper-small-jp-ONNX
whisper-small-jp-ONNX is a automatic speech recognition model from willopcbeta. Use it when you need speech turned into text. It is set up for transformers.js. The card lists the license as apache-2.0.
This is an ONNX version of drepic/whisper-small-jp. It was automatically converted and uploaded using this Hugging Face Space.
Downloads · 30 days
22
19% of all-time downloads
All-time downloads
117
Public
Repo size
7 GB
Likes
0
Public
Click a slice to open those files.
.onnx9.9 GB · 100%
From the Hugging Face model README
This is an ONNX version of drepic/whisper-small-jp. It was automatically converted and uploaded using this Hugging Face Space.
The optimal Q4 quantitative configuration: using decoder_model with q4f16 results in less ambiguous or nonsensical outputs.
quantization: {
encoder_model: 'q4f16',
decoder_model_merged: 'q4',
},
See the pipeline documentation for automatic-speech-recognition: https://huggingface.co/docs/transformers.js/api/pipelines#module_pipelines.AutomaticSpeechRecognitionPipeline
This model is a fine-tuned version of openai/whisper-small on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Wer | Cer |
|---|---|---|---|---|---|
| 0.6589 | 1.0 | 7154 | 0.6615 | 0.2735 | 0.2735 |
| 0.6273 | 2.0 | 14308 | 0.6457 | 0.2699 | 0.2699 |
| 0.6251 | 3.0 | 21462 | 0.6359 | 0.2660 | 0.2660 |
| 0.6427 | 4.0 | 28616 | 0.6283 | 0.2642 | 0.2642 |
| 0.6389 | 5.0 | 35770 | 0.6243 | 0.2631 | 0.2631 |
| 0.6078 | 6.0 | 42924 | 0.6242 | 0.2615 | 0.2615 |
| 0.5788 | 7.0 | 50078 | 0.6195 | 0.2603 | 0.2603 |
| 0.5801 | 8.0 | 57232 | 0.6180 | 0.2596 | 0.2596 |
| 0.5866 | 9.0 | 64386 | 0.6145 | 0.2598 | 0.2598 |
| 0.6052 | 10.0 | 71540 | 0.6168 | 0.2600 | 0.2600 |