Skip to content

mo-thecreator

ViT-GPT2-Image-Captioning

mo-thecreator/ViT-GPT2-Image-Captioning

ViT-GPT2-Image-Captioning is a image-to-text model from mo-thecreator. Use it when you need a caption or text from an image. It is set up for transformers.

should probably proofread and complete it, then remove this comment. --

Downloads · 30 days

15

2% of all-time downloads

All-time downloads

917

Public

Parameters

239M

1.9 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors957 MB · 100%

At a glance

Task
Image-to-Text
Library
transformers
Model type
vision-encoder-decoder
Access
Public
Created
Sep 30, 2024
Updated
Oct 7, 2024
SHA
73fc76f5

Base models

Task
Image-to-Text
Library
transformers
Type
vision-encoder-decoder
Created
Sep 30, 2024
Updated
Oct 7, 2024
ViT-GPT2-Image-Captioning — AI Model — AIMarketly