Skip to content

baseplate

vit-gpt2-image-captioning

baseplate/vit-gpt2-image-captioning

vit-gpt2-image-captioning is a image-to-text model from baseplate. Use it when you need a caption or text from an image. It is set up for transformers. The card lists the license as apache-2.0.

This is an image captioning model trained by @ydshieh in flax this is pytorch version of this.

Downloads · 30 days

91

34% of all-time downloads

All-time downloads

266

Public

Repo size

2 GB

Likes

2

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin982 MB · 100%

At a glance

Task
Image-to-Text
Library
transformers
License
apache-2.0
Model type
vision-encoder-decoder
Access
Public
Created
Apr 5, 2023
Updated
Apr 5, 2023
SHA
4ac3bf9d
Task
Image-to-Text
Library
transformers
Type
vision-encoder-decoder
License
apache-2.0
Created
Apr 5, 2023
Updated
Apr 5, 2023