Downloads · 30 days
119
2% of all-time downloads
ermu2001/pllava-7b
pllava-7b is a text generation model from ermu2001. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Model type: PLLaVA-7B is an open-source video-language chatbot trained by fine-tuning Image-LLM on video instruction-following data. It is an auto-regressive language model, based on the transformer architecture. Base…
Downloads · 30 days
119
2% of all-time downloads
All-time downloads
7.2K
Public
Parameters
7.1B
14.3 GB on disk
Likes
12
Public
Click a slice to open those files.
.safetensors14.3 GB · 100%
From the Hugging Face model README
Model type: PLLaVA-7B is an open-source video-language chatbot trained by fine-tuning Image-LLM on video instruction-following data. It is an auto-regressive language model, based on the transformer architecture. Base LLM: llava-hf/llava-v1.6-vicuna-7b-hf
Model date: PLLaVA-7B was trained in April 2024.
Paper or resources for more information:
llava-hf/llava-v1.6-vicuna-7b-hf license.
Where to send questions or comments about the model: https://github.com/magic-research/PLLaVA/issues
Primary intended uses: The primary use of PLLaVA is research on large multimodal models and chatbots.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
Video-Instruct-Tuning data of OpenGVLab/VideoChat2-IT
A collection of 6 benchmarks, including 5 VQA benchmarks and 1 recent benchmarks specifically proposed for Video-LMMs.