Skip to content

wangzhen-w

PanoVLN_base

wangzhen-w/PanoVLN_base

PanoVLN_base is a image-text-to-text model from wangzhen-w. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as other.

PanoVLN is a vision-and-language navigation policy that follows instructions from 360° RGB observations. It combines a Qwen3.5-4B backbone with PanoVGGT geometry features and predicts 18-action sequences. Confidence-g…

Downloads · 30 days

18

100% of all-time downloads

All-time downloads

18

Public

Parameters

6B

12 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors12 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
other
Model type
qwen3_5
Access
Public
Created
Sep 20, 2026
Updated
Oct 1, 2026
SHA
bb6f3121

Try a prompt

Task
Image-Text-to-Text
Library
transformers
Type
qwen3_5
License
other
Created
Sep 20, 2026
Updated
Oct 1, 2026
PanoVLN_base — AI Model — AIMarketly