Skip to content

aidansmyth95

nanoVLM

aidansmyth95/nanoVLM

nanoVLM is a image-text-to-text model from aidansmyth95. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for nanovlm. The card lists the license as mit.

nanoVLM is a minimal and lightweight Vision-Language Model (VLM) designed for efficient training and experimentation. Built using pure PyTorch, the entire model architecture and training logic fits within ~750 lines o…

Downloads · 30 days

17

31% of all-time downloads

All-time downloads

55

Public

Parameters

228M

2.7 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors912 MB · 100%

At a glance

Task
Image-Text-to-Text
Library
nanovlm
License
mit
Access
Public
Created
Oct 1, 2025
Updated
Jul 25, 2026
SHA
eac47f96

Try a prompt

Task
Image-Text-to-Text
Library
nanovlm
License
mit
Created
Oct 1, 2025
Updated
Jul 25, 2026