Downloads · 30 days
25
100% of all-time downloads
BlackCatRoboticsAI/smolvla-base
smolvla-base is a machine learning model from BlackCatRoboticsAI. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for lerobot. The card lists the license as apache-2.0.
SmolVLA — A compact Vision-Language-Action model from Hugging Face for robot manipulation. Takes images + language instructions, outputs robot actions.
Downloads · 30 days
25
100% of all-time downloads
All-time downloads
25
Public
Parameters
450M
907 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors907 MB · 100%
How the weights are stored.
BF16447M · 99%
From the Hugging Face model README
SmolVLA — A compact Vision-Language-Action model from Hugging Face for robot manipulation. Takes images + language instructions, outputs robot actions.
| Field | Value |
|---|---|
| License | Apache 2.0 (commercial use ✅) |
| Input | Multi-view images (256×256) + language instruction + proprioception |
| Output | Robot actions |
| Framework | LeRobot / PyTorch |
| Format | Safetensors |
from huggingface_hub import snapshot_download
from lerobot.policies.smolvla.modeling_smolvla import SmolVLA
# Download model
snapshot_download(repo_id="BlackCatRoboticsAI/smolvla-base", local_dir="./smolvla")
# Load policy
policy = SmolVLA.from_pretrained("./smolvla")
@misc{lerobot2025smolvla,
title={SmolVLA: Small Vision-Language-Action Model},
author={Hugging Face},
year={2025},
publisher={Hugging Face}
}