Skip to content

zhibinlan

LLaVE-0.5B

zhibinlan/LLaVE-0.5B

LLaVE-0.5B is a image-text-to-text model from zhibinlan. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.

The LLaVE models are 0.5B parameter multimodal embedding models based on the LLaVA-OneVision-0.5B model with a context window of 4K tokens.

Downloads · 30 days

23

0% of all-time downloads

All-time downloads

49.9K

Public

Parameters

894M

1.8 GB on disk

Likes

7

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors1.8 GB · 99%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
apache-2.0
Model type
llava
Access
Public
Created
Feb 6, 2025
Updated
Mar 14, 2025
SHA
a4769da3

Try a prompt

Base models

Task
Image-Text-to-Text
Library
transformers
Type
llava
License
apache-2.0
Languages
en
Created
Feb 6, 2025
Updated
Mar 14, 2025