Downloads · 30 days
12
11% of all-time downloads
woodwardmw/phase2
phase2 is a text-to-audio model from woodwardmw. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
11% of all-time downloads
All-time downloads
111
Public
Parameters
144M
12.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors578 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of microsoft/speecht5_tts on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.5917 | 2.6823 | 1000 | 0.5957 |
| 0.5782 | 5.3627 | 2000 | 0.5542 |
| 0.5593 | 8.0430 | 3000 | 0.5318 |
| 0.4973 | 10.7253 | 4000 | 0.4915 |
| 0.4883 | 13.4056 | 5000 | 0.4819 |
| 0.4868 | 16.0860 | 6000 | 0.4706 |
| 0.5138 | 18.7683 | 7000 | 0.4689 |
| 0.4571 | 21.4486 | 8000 | 0.4650 |
| 0.4557 | 24.1289 | 9000 | 0.4675 |
| 0.5072 | 26.8113 | 10000 | 0.4631 |
| 0.492 | 29.4916 | 11000 | 0.4604 |
| 0.4535 | 32.1719 | 12000 | 0.4581 |
| 0.4668 | 34.8543 | 13000 | 0.4550 |
| 0.4825 | 37.5346 | 14000 | 0.4593 |
| 0.4551 | 40.2149 | 15000 | 0.4568 |
| 0.4285 | 42.8972 | 16000 | 0.4554 |
| 0.4383 | 45.5776 | 17000 | 0.4544 |
| 0.393 | 48.2579 | 18000 | 0.4529 |
| 0.4406 | 50.9402 | 19000 | 0.4570 |
| 0.4519 | 53.6206 | 20000 | 0.4544 |