Skip to content

google

vit-base-patch32-224-in21k

google/vit-base-patch32-224-in21k

vit-base-patch32-224-in21k is a image feature extraction model from google. Use it for the image feature extraction task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.

Vision Transformer (ViT) model pre-trained on ImageNet-21k (14 million images, 21,843 classes) at resolution 224x224. It was introduced in the paper An Image is Worth 16x16 Words: Transformers for Image Recognition at…

Downloads · 30 days

22.5K

0% of all-time downloads

All-time downloads

6.8M

Public

Parameters

88M

1.8 GB on disk

Likes

20

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.h5352 MB · 25%

At a glance

Task
Image Feature Extraction
Library
transformers
License
apache-2.0
Model type
vit
Access
Public
Created
Mar 2, 2022
Updated
Dec 8, 2022
SHA
98cb6f70

Datasets

Task
Image Feature Extraction
Library
transformers
Type
vit
License
apache-2.0
Created
Mar 2, 2022
Updated
Dec 8, 2022