Skip to content

mo-thecreator

ViT-GPT2-Image_Captioning_model

mo-thecreator/ViT-GPT2-Image_Captioning_model

ViT-GPT2-Image_Captioning_model is a image-to-text model from mo-thecreator. Use it when you need a caption or text from an image. It is set up for transformers.

should probably proofread and complete it, then remove this comment. --

Downloads · 30 days

19

4% of all-time downloads

All-time downloads

462

Public

Parameters

239M

957 MB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors957 MB · 100%

At a glance

Task
Image-to-Text
Library
transformers
Model type
vision-encoder-decoder
Access
Public
Created
Sep 29, 2024
Updated
Sep 30, 2024
SHA
cb6dbf1c
Task
Image-to-Text
Library
transformers
Type
vision-encoder-decoder
Created
Sep 29, 2024
Updated
Sep 30, 2024
ViT-GPT2-Image_Captioning_model — AI Model — AIMarketly