Downloads · 30 days
17
7% of all-time downloads
Xenova/pix2struct-widget-captioning-base
pix2struct-widget-captioning-base is a image-text-to-text model from Xenova. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.js.
https://huggingface.co/google/pix2struct-widget-captioning-base with ONNX weights to be compatible with Transformers.js.
Downloads · 30 days
17
7% of all-time downloads
All-time downloads
233
Public
Repo size
3.8 GB
Likes
1
Public
Click a slice to open those files.
.onnx3.8 GB · 100%
From the Hugging Face model README
https://huggingface.co/google/pix2struct-widget-captioning-base with ONNX weights to be compatible with Transformers.js.
Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using 🤗 Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).