Downloads · 30 days
0
0% of all-time downloads
UCSC-VLAA/openvision-vit-huge-patch14-84
openvision-vit-huge-patch14-84 is a image feature extraction model from UCSC-VLAA. Use it for the image feature extraction task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
This repository contains the model weights based on the work described in OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning.
Downloads · 30 days
0
0% of all-time downloads
All-time downloads
19
Public
Repo size
15.4 GB
Likes
0
Public
Click a slice to open those files.
Other11.5 GB · 75%
From the Hugging Face model README
This repository contains the model weights based on the work described in OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning.
Project Page: https://ucsc-vlaa.github.io/OpenVision/
For details on training and usage, please refer to the Github repository.