Skip to content

nagayama0706

multimodal_model

nagayama0706/multimodal_model

multimodal_model is a visual question answering model from nagayama0706. Use it for the visual question answering task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.

multimodalmodel is a merge of the following models using LazyMergekit: CIDAS/clipseg-rd64-refined dalle-mini/dalle-mini

Downloads · 30 days

17

8% of all-time downloads

All-time downloads

222

Public

Parameters

7.2B

14.5 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors14.5 GB · 100%

At a glance

Task
Visual Question Answering
Library
transformers
License
apache-2.0
Model type
mistral
Access
Public
Created
Apr 16, 2024
Updated
Apr 16, 2024
SHA
abc86e3d

Base models

Task
Visual Question Answering
Library
transformers
Type
mistral
License
apache-2.0
Created
Apr 16, 2024
Updated
Apr 16, 2024
multimodal_model — AI Model — AIMarketly