Downloads · 30 days
5
29% of all-time downloads
Skyler215/SwinV2_Syllable
SwinV2_Syllable is a image-text-to-text model from Skyler215. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
29% of all-time downloads
All-time downloads
17
Public
Parameters
440M
15.9 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.8 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Bleu-4 |
|---|---|---|---|---|
| No log | 1.0 | 118 | 1.4441 | 0.1228 |
| No log | 2.0 | 236 | 1.3026 | 0.1406 |
| 2.0335 | 3.0 | 354 | 1.2238 | 0.1685 |
| 2.0335 | 4.0 | 472 | 1.1742 | 0.1841 |
| 2.0335 | 5.0 | 590 | 1.1522 | 0.1944 |
| 1.0795 | 6.0 | 708 | 1.1471 | 0.2103 |
| 1.0795 | 7.0 | 826 | 1.1434 | 0.2125 |
| 0.8075 | 8.0 | 944 | 1.1574 | 0.2108 |
| 0.8075 | 9.0 | 1062 | 1.1881 | 0.2121 |