Downloads · 30 days
1
5% of all-time downloads
Devishri1/finetuned_speecht5_50Kdata
finetuned_speecht5_50Kdata is a text-to-audio model from Devishri1. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
1
5% of all-time downloads
All-time downloads
20
Public
Parameters
144M
49.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.5055 | 0.3505 | 500 | 0.4585 |
| 0.4679 | 0.7010 | 1000 | 0.4248 |
| 0.4444 | 1.0515 | 1500 | 0.4140 |
| 0.4356 | 1.4020 | 2000 | 0.4109 |
| 0.434 | 1.7525 | 2500 | 0.4082 |
| 0.428 | 2.1030 | 3000 | 0.4023 |
| 0.423 | 2.4536 | 3500 | 0.4003 |
| 0.4177 | 2.8041 | 4000 | 0.3835 |
| 0.4077 | 3.1546 | 4500 | 0.3766 |
| 0.4096 | 3.5051 | 5000 | 0.3781 |