Downloads · 30 days
167
5% of all-time downloads
lmms-lab/LLaVA-NeXT-Video-34B
LLaVA-NeXT-Video-34B is a text generation model from lmms-lab. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Downloads · 30 days
167
5% of all-time downloads
All-time downloads
3.1K
Public
Parameters
34.8B
69.5 GB on disk
Likes
16
Public
Click a slice to open those files.
.safetensors69.5 GB · 100%
From the Hugging Face model README
Model type: <br> LLaVA-Next-Video is an open-source chatbot trained by fine-tuning LLM on multimodal instruction-following data. <br> Base LLM: NousResearch/Nous-Hermes-2-Yi-34B
Model date: <br> LLaVA-Next-Video-34B was trained in April 2024.
Paper or resources for more information: <br> https://github.com/LLaVA-VL/LLaVA-NeXT
NousResearch/Nous-Hermes-2-Yi-34B license.
https://github.com/LLaVA-VL/LLaVA-NeXT/issues
Primary intended uses: <br> The primary use of LLaVA is research on large multimodal models and chatbots.
Primary intended users: <br> The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
A collection of 4 benchmarks, including 3 academic VQA benchmarks and 1 captioning benchmark.