Downloads · 30 days
12
5% of all-time downloads
froabera/speecht5_finetuned
speecht5_finetuned is a text-to-audio model from froabera. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
5% of all-time downloads
All-time downloads
225
Public
Parameters
144M
19.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
<a href="https://huggingface.co/spaces/froabera/trackio" target="_blank"><img src="https://raw.githubusercontent.com/gradio-app/trackio/refs/heads/main/trackio/assets/badge.png" alt="Visualize in Trackio" title="Visualize in Trackio" style="height: 40px;"/></a>
This model is a fine-tuned version of microsoft/speecht5_tts on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 3.9966 | 4.7187 | 1000 | 0.4597 |
| 3.7975 | 9.4350 | 2000 | 0.4403 |
| 3.7564 | 14.1513 | 3000 | 0.4319 |
| 3.7496 | 18.8700 | 4000 | 0.4287 |