Downloads · 30 days
7
28% of all-time downloads
Sakura927/PyPE-LLaVA-7B
PyPE-LLaVA-7B is a image-text-to-text model from Sakura927. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
<h3PyPE: Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding</h3
Downloads · 30 days
7
28% of all-time downloads
All-time downloads
25
Public
Parameters
7.1B
14.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors14.1 GB · 100%
From the Hugging Face model README
For more details, please refer to Github: PyPE.