Downloads · 30 days
5
3% of all-time downloads
rahafvii/HopefulASR
HopefulASR is a text-to-audio model from rahafvii. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
3% of all-time downloads
All-time downloads
143
Public
Parameters
144M
11 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.4624 | 3.0769 | 100 | 0.4358 |
| 0.4523 | 6.1538 | 200 | 0.4288 |
| 0.4385 | 9.2308 | 300 | 0.4148 |
| 0.4288 | 12.3077 | 400 | 0.4126 |
| 0.4262 | 15.3846 | 500 | 0.4114 |
| 0.4236 | 18.4615 | 600 | 0.4128 |
| 0.416 | 21.5385 | 700 | 0.4105 |
| 0.4118 | 24.6154 | 800 | 0.4111 |
| 0.4091 | 27.6923 | 900 | 0.4074 |
| 0.4044 | 30.7692 | 1000 | 0.4085 |
| 0.4048 | 33.8462 | 1100 | 0.4080 |
| 0.402 | 36.9231 | 1200 | 0.4021 |
| 0.3957 | 40.0 | 1300 | 0.4034 |
| 0.3977 | 43.0769 | 1400 | 0.4027 |
| 0.3982 | 46.1538 | 1500 | 0.4041 |
| 0.395 | 49.2308 | 1600 | 0.4003 |
| 0.3913 | 52.3077 | 1700 | 0.4017 |
| 0.3969 | 55.3846 | 1800 | 0.4032 |
| 0.3918 | 58.4615 | 1900 | 0.4024 |
| 0.3898 | 61.5385 | 2000 | 0.4022 |