Downloads · 30 days
4
2% of all-time downloads
SitholeDavid/speecht5_improved_data_wo_knn
speecht5_improved_data_wo_knn is a text-to-audio model from SitholeDavid. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
4
2% of all-time downloads
All-time downloads
208
Public
Parameters
144M
1.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.9419 | 0.8097 | 250 | 0.8022 |
| 0.7827 | 1.6194 | 500 | 0.6917 |
| 0.7165 | 2.4291 | 750 | 0.6538 |
| 0.7038 | 3.2389 | 1000 | 0.6365 |
| 0.6821 | 4.0486 | 1250 | 0.6194 |
| 0.685 | 4.8583 | 1500 | 0.6156 |
| 0.6608 | 5.6680 | 1750 | 0.6083 |
| 0.6725 | 6.4777 | 2000 | 0.6051 |
| 0.6592 | 7.2874 | 2250 | 0.6009 |
| 0.6617 | 8.0972 | 2500 | 0.5981 |
| 0.6565 | 8.9069 | 2750 | 0.5962 |
| 0.6536 | 9.7166 | 3000 | 0.5957 |