Downloads · 30 days
56
10% of all-time downloads
OpenGVLab/Docopilot-2B
Docopilot-2B is a image-text-to-text model from OpenGVLab. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
This model is in the paper Docopilot: Improving Multimodal Models for Document-Level Understanding.
Downloads · 30 days
56
10% of all-time downloads
All-time downloads
586
Public
Parameters
2.2B
4.4 GB on disk
Likes
8
Public
Click a slice to open those files.
.safetensors4.4 GB · 100%
From the Hugging Face model README
This model is in the paper Docopilot: Improving Multimodal Models for Document-Level Understanding.
Please refer to https://github.com/OpenGVLab/Docopilot for details.