Skip to content

divinetribe

Vision-Narrator-0.8B-4bit-mlx

divinetribe/Vision-Narrator-0.8B-4bit-mlx

Vision-Narrator-0.8B-4bit-mlx is a image-text-to-text model from divinetribe. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for mlx. The card lists the license as apache-2.0.

A vision-language model small enough to live on a phone, trained to answer the questions blind people actually ask about what's in front of them.

Downloads · 30 days

90

76% of all-time downloads

All-time downloads

119

Public

Parameters

853M

1.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors625 MB · 96%

Parameter types

How the weights are stored.

U32752M · 88%

Try a prompt

Base models

Task
Image-Text-to-Text
Library
mlx
Type
qwen3_5
License
apache-2.0
Languages
en
Created
Sep 8, 2026
Updated
Sep 22, 2026