Skip to content

tifa-benchmark

promptcap-coco-vqa

tifa-benchmark/promptcap-coco-vqa

promptcap-coco-vqa is a image-to-text model from tifa-benchmark. Use it when you need a caption or text from an image. It is set up for transformers. The card lists the license as openrail.

This is the repo for the paper PromptCap: Prompt-Guided Task-Aware Image Captioning. This paper is accepted to ICCV 2023 as PromptCap: Prompt-Guided Image Captioning for VQA with GPT-3.

Downloads · 30 days

25

0% of all-time downloads

All-time downloads

85.8K

Public

Repo size

4.9 GB

Likes

15

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin2.4 GB · 100%

At a glance

Task
Image-to-Text
Library
transformers
License
openrail
Model type
ofa
Access
Public
Created
Jan 23, 2023
Updated
Dec 11, 2023
SHA
de7d203d
Task
Image-to-Text
Library
transformers
Type
ofa
License
openrail
Languages
en
Created
Jan 23, 2023
Updated
Dec 11, 2023
promptcap-coco-vqa — AI Model — AIMarketly