Downloads · 30 days
1.8K
4% of all-time downloads
mlx-community/SmolVLM2-500M-Video-Instruct-mlx
SmolVLM2-500M-Video-Instruct-mlx is a video-text-to-text model from mlx-community. Use it for the video-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
This model was converted to MLX format from HuggingFaceTB/SmolVLM2-500M-Video-Instruct using mlx-vlm version 0.1.13. Refer to the original model card for more details on the model.
Downloads · 30 days
1.8K
4% of all-time downloads
All-time downloads
45.3K
Public
Parameters
507M
1 GB on disk
Likes
18
Public
Click a slice to open those files.
.safetensors1 GB · 100%
From the Hugging Face model README
This model was converted to MLX format from HuggingFaceTB/SmolVLM2-500M-Video-Instruct using mlx-vlm version 0.1.13.
Refer to the original model card for more details on the model.
pip install -U mlx-vlm
python -m mlx_vlm.generate --model mlx-community/SmolVLM2-500M-Video-Instruct-mlx --image https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/bee.jpg --prompt "Can you describe this image?"