Downloads · 30 days
12
23% of all-time downloads
larsvandorp/folding_pi05
folding_pi05 is a robotics model from larsvandorp. Use it for the robotics task on the model card, and read the license before you ship it in a product. It is set up for lerobot. The card lists the license as apache-2.0.
π0.5 vision-language-action policy fine-tuned for autonomous towel folding on a 6-DoF SO-101 follower arm with a single wrist camera. A strong alternative to the diffusion-transformer policy larsvandorp/foldingdit.
Downloads · 30 days
12
23% of all-time downloads
All-time downloads
53
Public
Parameters
4.1B
9.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors9.4 GB · 100%
How the weights are stored.
BF163.6B · 87%
From the Hugging Face model README
π0.5 vision-language-action policy fine-tuned for autonomous towel folding on a 6-DoF SO-101 follower arm with a single wrist camera. A strong alternative to the diffusion-transformer policy larsvandorp/folding_dit.
The repo root holds the step-9000 checkpoint (our best), so from_pretrained("larsvandorp/folding_pi05") loads it directly.
lerobot/pi05_base (PaliGemma backbone + action expert, ~3B params), vision encoder unfrozen.uv venv --python 3.12 .venv
GIT_LFS_SKIP_SMUDGE=1 uv pip install --python .venv/bin/python \
"lerobot[pi] @ git+https://github.com/LarsvanDorp/lerobot.git@dinov3"
.venv/bin/lerobot-rollout \
--strategy.type=base \
--robot.type=so101_follower --robot.port=/dev/ttyACM0 --robot.id=my_follower \
--robot.cameras="{wrist: {type: opencv, index_or_path: <cam-index>, width: 800, height: 600, fps: 30, fourcc: MJPG}}" \
--policy.path=larsvandorp/folding_pi05 \
--policy.device=cuda --inference.type=sync \
--task="fold the towel" --duration=60
Note the fourcc: MJPG in the camera config (needed on the lab Linux PC). We run without --interpolation_multiplier.
larsvandorp/magic_soup — the filtered SO-101 towel-folding set (bad episodes removed: high mean |Δa|, or no fold in the last frame).