Downloads · 30 days
3
1% of all-time downloads
Kishor798/speecht5_finetuned_sample
speecht5_finetuned_sample is a text-to-audio model from Kishor798. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
3
1% of all-time downloads
All-time downloads
224
Public
Parameters
144M
36.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.581 | 88.8889 | 100 | 0.6059 |
| 0.4879 | 177.7778 | 200 | 0.5648 |
| 0.4272 | 266.6667 | 300 | 0.5928 |
| 0.3992 | 355.5556 | 400 | 0.6150 |
| 0.3642 | 444.4444 | 500 | 0.5903 |
| 0.349 | 533.3333 | 600 | 0.6436 |
| 0.3298 | 622.2222 | 700 | 0.6391 |
| 0.3178 | 711.1111 | 800 | 0.6306 |
| 0.3041 | 800.0 | 900 | 0.6405 |
| 0.2981 | 888.8889 | 1000 | 0.6490 |
| 0.2913 | 977.7778 | 1100 | 0.6502 |
| 0.2846 | 1066.6667 | 1200 | 0.6444 |
| 0.2873 | 1155.5556 | 1300 | 0.6636 |
| 0.2857 | 1244.4444 | 1400 | 0.6504 |
| 0.2839 | 1333.3333 | 1500 | 0.6517 |