Skip to content

ilpa-user

mp3_nc_vision

ilpa-user/mp3_nc_vision

mp3_nc_vision is a image-text-to-text model from ilpa-user. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as llama3.2.

The Llama 3.2-Vision collection of multimodal large language models (LLMs) is a collection of pretrained and instruction-tuned image reasoning generative models in 11B and 90B sizes (text \+ images in / text out). The…

Downloads · 30 days

4

22% of all-time downloads

All-time downloads

18

Public

Repo size

21.5 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors21.4 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
llama3.2
Model type
mllama
Access
Public
Created
Feb 3, 2025
Updated
Feb 3, 2025
SHA
77b53186
Task
Image-Text-to-Text
Library
transformers
Type
mllama
License
llama3.2
Languages
en, de, fr, it, pt, hi
Created
Feb 3, 2025
Updated
Feb 3, 2025
mp3_nc_vision — AI Model — AIMarketly