Downloads · 30 days
2
13% of all-time downloads
binisha/speecht5_finetune_binisha
speecht5_finetune_binisha is a text-to-audio model from binisha. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
2
13% of all-time downloads
All-time downloads
16
Public
Parameters
144M
17.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.6028 | 2.7586 | 100 | 0.5187 |
| 0.5195 | 5.5172 | 200 | 0.4851 |
| 0.5075 | 8.2759 | 300 | 0.4708 |
| 0.462 | 11.0345 | 400 | 0.4609 |
| 0.4429 | 13.7931 | 500 | 0.4294 |
| 0.4303 | 16.5517 | 600 | 0.4249 |
| 0.4172 | 19.3103 | 700 | 0.4184 |
| 0.402 | 22.0690 | 800 | 0.4077 |
| 0.3898 | 24.8276 | 900 | 0.3975 |
| 0.3966 | 27.5862 | 1000 | 0.4197 |
| 0.3773 | 30.3448 | 1100 | 0.3955 |
| 0.3658 | 33.1034 | 1200 | 0.3878 |
| 0.3644 | 35.8621 | 1300 | 0.3878 |
| 0.3622 | 38.6207 | 1400 | 0.3841 |
| 0.3671 | 41.3793 | 1500 | 0.3836 |