Downloads · 30 days
11
46% of all-time downloads
VINHYU/OpenSpatial-InternVL2.5-8B
OpenSpatial-InternVL2.5-8B is a image-text-to-text model from VINHYU. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
OpenSpatial-InternVL2.5-8B is an OpenSpatial model fine-tuned for spatial understanding and reasoning. This repository contains full, inference-ready model weights.
Downloads · 30 days
11
46% of all-time downloads
All-time downloads
24
Public
Parameters
8.1B
16.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.2 GB · 100%
From the Hugging Face model README
OpenSpatial-InternVL2.5-8B is an OpenSpatial model fine-tuned for spatial understanding and reasoning. This repository contains full, inference-ready model weights.
Use transformers the version recommended by InternVL2.5 for this checkpoint.
from transformers import AutoModel, AutoTokenizer
model_id = "VINHYU/OpenSpatial-InternVL2.5-8B"
model = AutoModel.from_pretrained(
model_id,
torch_dtype="auto",
low_cpu_mem_usage=True,
trust_remote_code=True,
).eval().cuda()
tokenizer = AutoTokenizer.from_pretrained(
model_id, trust_remote_code=True, use_fast=False
)
Please refer to the OpenSpatial repository and the base model repository for full usage instructions, limitations, and license terms.