Downloads · 30 days
19
4% of all-time downloads
mo-thecreator/ViT-GPT2-Image_Captioning_model
ViT-GPT2-Image_Captioning_model is a image-to-text model from mo-thecreator. Use it when you need a caption or text from an image. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
19
4% of all-time downloads
All-time downloads
462
Public
Parameters
239M
957 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors957 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Rouge2 Fmeasure |
|---|---|---|---|---|
| No log | 0.9987 | 496 | 2.4901 | 0.1077 |
| 2.5089 | 1.9995 | 993 | 2.4292 | 0.1141 |
| 2.4103 | 2.9962 | 1488 | 2.4134 | 0.1166 |