Downloads · 30 days
20
23% of all-time downloads
L0Xit/KANIME-V1
KANIME-V1 is a image-text-to-text model from L0Xit. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
This repository contains the MangaLMM model described in the paper MangaVQA and MangaLMM: A Benchmark and Specialized Model for Multimodal Manga Understanding.
Downloads · 30 days
20
23% of all-time downloads
All-time downloads
86
Public
Parameters
8.3B
16.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.6 GB · 100%
From the Hugging Face model README
This repository contains the MangaLMM model described in the paper MangaVQA and MangaLMM: A Benchmark and Specialized Model for Multimodal Manga Understanding.
Code: https://github.com/manga109/MangaLMM <br> Official demo: https://huggingface.co/spaces/yuki-imajuku/MangaLMM-Demo