Downloads · 30 days
80
6% of all-time downloads
ermu2001/pllava-34b
pllava-34b is a text generation model from ermu2001. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Model type: PLLaVA-34B is an open-source video-language chatbot trained by fine-tuning Image-LLM on video instruction-following data. It is an auto-regressive language model, based on the transformer architecture. Bas…
Downloads · 30 days
80
6% of all-time downloads
All-time downloads
1.3K
Public
Parameters
34.9B
69.9 GB on disk
Likes
14
Public
Click a slice to open those files.
.safetensors69.9 GB · 100%
From the Hugging Face model README
Model type: PLLaVA-34B is an open-source video-language chatbot trained by fine-tuning Image-LLM on video instruction-following data. It is an auto-regressive language model, based on the transformer architecture. Base LLM: liuhaotian/llava-v1.6-34b
Model date: PLLaVA-34B was trained in April 2024.
Paper or resources for more information:
NousResearch/Nous-Hermes-2-Yi-34B license.
Where to send questions or comments about the model: https://github.com/magic-research/PLLaVA/issues
Primary intended uses: The primary use of PLLaVA is research on large multimodal models and chatbots.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
Video-Instruct-Tuning data of OpenGVLab/VideoChat2-IT
A collection of 6 benchmarks, including 5 Video QA benchmarks and 1 benchmarks specifically proposed for Video-LMMs.