Skip to content

introvoyz041

MiniVLA

introvoyz041/MiniVLA

MiniVLA is a image-text-to-text model from introvoyz041. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.

This repository hosts MiniVLA – a modular and deployment-friendly Vision-Language-Action (VLA) model designed for edge hardware (e.g., Jetson Orin Nano). It contains model checkpoints, Hugging Face–compatible Qwen-0.5…

Downloads · 30 days

0

Access

Public

Updated May 28, 2026

Repo size

12.9 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt5.6 GB · 43%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
apache-2.0
Access
Public
Created
May 28, 2026
Updated
May 28, 2026
SHA
92b27168

Try a prompt

Base models

Datasets

Task
Image-Text-to-Text
Library
transformers
License
apache-2.0
Languages
en
Created
May 28, 2026
Updated
May 28, 2026