Downloads · 30 days
5
3% of all-time downloads
statking/paligemma_vqa_lower
paligemma_vqa_lower is a image-text-to-text model from statking. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as gemma.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
3% of all-time downloads
All-time downloads
185
Public
Parameters
2.9B
10.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.8 GB · 100%
From the Hugging Face model README
This model is a fine-tuned version of google/paligemma-3b-pt-224 on the vq_av2 dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 4.837 | 0.1471 | 500 | 3.7992 |
| 0.1673 | 0.2943 | 1000 | 0.1149 |
| 0.0227 | 0.4414 | 1500 | 0.0198 |
| 0.0146 | 0.5886 | 2000 | 0.0138 |
| 0.0135 | 0.7357 | 2500 | 0.0125 |
| 0.013 | 0.8829 | 3000 | 0.0122 |