Downloads · 30 days
2
13% of all-time downloads
Mohsen21/ESP_TWICE_DATA
ESP_TWICE_DATA is a text-to-audio model from Mohsen21. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
2
13% of all-time downloads
All-time downloads
16
Public
Parameters
144M
8.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.5685 | 5.1613 | 100 | 0.5270 |
| 0.5244 | 10.3226 | 200 | 0.4922 |
| 0.5034 | 15.4839 | 300 | 0.4847 |
| 0.4875 | 20.6452 | 400 | 0.4703 |
| 0.4759 | 25.8065 | 500 | 0.4669 |
| 0.4718 | 30.9677 | 600 | 0.4664 |
| 0.4604 | 36.1290 | 700 | 0.4609 |
| 0.4526 | 41.2903 | 800 | 0.4637 |
| 0.4525 | 46.4516 | 900 | 0.4602 |
| 0.4452 | 51.6129 | 1000 | 0.4608 |
| 0.448 | 56.7742 | 1100 | 0.4622 |
| 0.4373 | 61.9355 | 1200 | 0.4580 |
| 0.4326 | 67.0968 | 1300 | 0.4580 |
| 0.4335 | 72.2581 | 1400 | 0.4594 |
| 0.4328 | 77.4194 | 1500 | 0.4606 |