Downloads · 30 days
23
2% of all-time downloads
UCSC-VLAA/openvision-vit-base-patch16-224
openvision-vit-base-patch16-224 is a image feature extraction model from UCSC-VLAA. Use it for the image feature extraction task on the model card, and read the license before you ship it in a product. It is set up for open_clip. The card lists the license as apache-2.0.
This repository contains the OpenVision ViT model described in the paper OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning.
Downloads · 30 days
23
2% of all-time downloads
All-time downloads
1.2K
Public
Repo size
2.3 GB
Likes
0
Public
Click a slice to open those files.
Other1.7 GB · 75%
From the Hugging Face model README
This repository contains the OpenVision ViT model described in the paper OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning.
Project page: https://ucsc-vlaa.github.io/OpenVision