Downloads · 30 days
0
sanatem/samtp-mini-traversability
samtp-mini-traversability is a robotics model from sanatem. Use it for the robotics task on the model card, and read the license before you ship it in a product. It is set up for sam-tp. The card lists the license as apache-2.0.
Image-space traversability segmentation for the FrodoBots Earth Rover Mini+ front camera. One RGB frame in, per-pixel drivability out.
Downloads · 30 days
0
Access
Public
Updated Aug 6, 2026
Repo size
137 MB
Likes
0
Public
Click a slice to open those files.
.pt137 MB · 100%
From the Hugging Face model README
Image-space traversability segmentation for the FrodoBots Earth Rover Mini+ front camera. One RGB frame in, per-pixel drivability out.
CustomPromptEncoderLarger, want_custom_prompt_encoder: 2).
Prompt-free: point/box inputs are ignored; output is deterministic per image.facebook/sam2.1-hiera-tinycheckpoint_finetuned_v2.pt — torch save with a single top-level
model key holding the state dict. 136,622,641 bytes.44e508da3d36a63431f8197f16784c980abf43ea94fc4e524bcd19d0646692bdThis checkpoint ONLY loads against the tiny SAM-TP inference config
(sam2/configs/sam2.1_inference_tiny/sam2.1_custom2.yaml in the GeNIE sam2
fork). Loading it with a base+/small/large config — or loading the public GeNIE
checkpoint_2.pt (base+) with the tiny config — fails with a state-dict
mismatch.
Via the rover-traversability package (in the team repo under traversability/):
pip install 'rover-traversability[hf]' # from the repo: pip install -e ./traversability[hf]
python -c "
from rover_traversability import TraversabilityPredictor
p = TraversabilityPredictor() # auto-downloads this checkpoint
result = p.predict('frame.jpg')
print(result.mask.shape, result.mask.mean())
"
Or manually: download checkpoint_finetuned_v2.pt and set
SAMTP_CHECKPOINT=/path/to/checkpoint_finetuned_v2.pt.
Output contract: mask is HxW float32 in [0, 1], 1 = drivable (sigmoid of the
raw logits, resized to the input frame size).
This is a full model state dict — use it directly as the init checkpoint in
Meta's SAM2 training harness (training.* from facebookresearch/sam2) with the
sam2.1_training_tiny configs from the GeNIE fork (ckpt_state_dict_keys: ['model']). Dataset format: image folder + binary PNG masks (MOSE/PNG-VOS
layout). Reference hyperparameters from this checkpoint's training: 1024 res,
batch 8, AdamW, base_lr 5e-6 / vision_lr 3e-6, 5 epochs.
| Device | Latency/frame |
|---|---|
| MPS | ~0.16–0.23 s (4–6 Hz) |
| CPU | ~0.44 s (~2.3 Hz) |
rover-traversability wrapper applies a per-frame luminance-contrast
refinement to mitigate this — keep it enabled.BitRobot/FrodoBots-Mini-4K
(CC-BY-SA) — if you redistribute or build on these weights, carry this
provenance note and attribution with them.