Skip to content

nectec

Pathumma-llm-vision-1.0.0

nectec/Pathumma-llm-vision-1.0.0

Pathumma-llm-vision-1.0.0 is a visual question answering model from nectec. Use it for the visual question answering task on the model card, and read the license before you ship it in a product.

Pathumma-llm-vision-1.0.0 is a multi-modal language model fine-tuned for Visual Question Answering (VQA) and Image Captioning tasks. It contains 8 billion parameters and leverages both image and text processing to und…

Downloads · 30 days

18

1% of all-time downloads

All-time downloads

1.7K

Public

Parameters

8.5B

40.4 GB on disk

Likes

11

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors30.9 GB · 100%

Parameter types

How the weights are stored.

F327B · 82%

Base models

Task
Visual Question Answering
Type
idefics3
Languages
th, en
Created
Oct 24, 2024
Updated
Oct 25, 2024