Downloads · 30 days
3
21% of all-time downloads
Wsassi/speecht5_finetuned_polyAi
speecht5_finetuned_polyAi is a text-to-audio model from Wsassi. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
3
21% of all-time downloads
All-time downloads
14
Public
Parameters
144M
578 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on the minds14 dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 1.6267 | 1.0 | 124 | 1.0811 |
| 0.9903 | 2.0 | 248 | 0.7723 |
| 0.95 | 3.0 | 372 | 0.7042 |
| 0.8184 | 4.0 | 496 | 0.6570 |
| 0.7998 | 5.0 | 620 | 0.6495 |