Downloads · 30 days
0
AsadIsmail/SmolVLM2-500M-Video-Instruct-ternary
SmolVLM2-500M-Video-Instruct-ternary is a image-text-to-text model from AsadIsmail. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
Ternary-quantized version of HuggingFaceTB/SmolVLM2-500M-Video-Instruct — a compact video-understanding VLM.
Downloads · 30 days
0
Access
Public
Updated Apr 17, 2026
Repo size
368 MB
Likes
0
Public
Click a slice to open those files.
.npz368 MB · 100%
From the Hugging Face model README
Ternary-quantized version of HuggingFaceTB/SmolVLM2-500M-Video-Instruct — a compact video-understanding VLM.
| Property | Value |
|---|---|
| Base Model | HuggingFaceTB/SmolVLM2-500M-Video-Instruct |
| Parameters | 500M |
| Quantization | tritplane3 (225 linear layers) |
| Full-model effective bits | 11.50 |
| Compression ratio | 1.39× |
| Avg reconstruction error | 0.1413 |
| Vision Encoder | FP16 (preserved) |
Part of ternary-models.