Skip to content

cminst

Llama-3.2-11B-VisionEncoder

cminst/Llama-3.2-11B-VisionEncoder

Llama-3.2-11B-VisionEncoder is a image feature extraction model from cminst. Use it for the image feature extraction task on the model card, and read the license before you ship it in a product. It is set up for transformers.

This repository contains only the MllamaVisionModel weights extracted from unsloth/Llama-3.2-11B-Vision. It intentionally excludes the text decoder and language model weights.

Downloads · 30 days

28

10% of all-time downloads

All-time downloads

279

Public

Parameters

836M

3.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors3.3 GB · 100%

At a glance

Task
Image Feature Extraction
Library
transformers
Model type
mllama_vision_model
Access
Public
Created
Jun 8, 2026
Updated
Jun 8, 2026
SHA
621808b2

Base models

Task
Image Feature Extraction
Library
transformers
Type
mllama_vision_model
Created
Jun 8, 2026
Updated
Jun 8, 2026
Llama-3.2-11B-VisionEncoder — AI Model — AIMarketly