Downloads · 30 days
12
17% of all-time downloads
lithish2602/speecht5_tts_ta
speecht5_tts_ta is a text-to-audio model from lithish2602. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
17% of all-time downloads
All-time downloads
70
Public
Parameters
144M
7.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on the common_voice_17_0 dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.3035 | 500.0 | 1000 | 0.6257 |
| 0.2769 | 1000.0 | 2000 | 0.6513 |
| 0.2608 | 1500.0 | 3000 | 0.6888 |
| 0.2503 | 2000.0 | 4000 | 0.6856 |