Downloads · 30 days
15
1% of all-time downloads
cs-giung/clip-vit-base-patch16-laion2b
clip-vit-base-patch16-laion2b is a zero-shot image classification model from cs-giung. Use it for the zero-shot image classification task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
Contrastive Language-Image Pretraining (CLIP) model pre-trained on LAION-2B at resolution 224x224. It was introduced in the paper Learning Transferable Visual Models From Natural Language Supervision and further repro…
Downloads · 30 days
15
1% of all-time downloads
All-time downloads
1.1K
Public
Parameters
150M
599 MB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors599 MB · 99%
How the weights are stored.
F32150M · 100%
From the Hugging Face model README
Contrastive Language-Image Pretraining (CLIP) model pre-trained on LAION-2B at resolution 224x224. It was introduced in the paper Learning Transferable Visual Models From Natural Language Supervision and further reproduced in the follow-up paper Reproducible scaling laws for contrastive language-image learning.
The weights were converted from the laion/CLIP-ViT-B-16-laion2B-s34B-b88K presented in the OpenCLIP LAION-2B collections.