Downloads · 30 days
5
12% of all-time downloads
itsnotacreativeuser/SmolVLM-256M-ScreenTask
SmolVLM-256M-ScreenTask is a image-text-to-text model from itsnotacreativeuser. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
5
12% of all-time downloads
All-time downloads
41
Public
Repo size
11.6 MB
Likes
0
Public
Click a slice to open those files.
.safetensors11.6 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of HuggingFaceTB/SmolVLM-256M-Instruct on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 2.857 | 0.0599 | 20 | 2.7762 |
| 1.8633 | 0.1199 | 40 | 1.8233 |
| 1.1164 | 0.1798 | 60 | 1.0640 |
| 0.8947 | 0.2397 | 80 | 0.8909 |
| 0.8471 | 0.2996 | 100 | 0.8402 |