Downloads · 30 days
19
7% of all-time downloads
xtuner/llava-phi-3-mini-pretrain
llava-phi-3-mini-pretrain is a visual question answering model from xtuner. Use it for the visual question answering task on the model card, and read the license before you ship it in a product. It is set up for transformers.
<div align="center" <img src="https://github.com/InternLM/lmdeploy/assets/36994684/0cf8d00f-e86b-40ba-9b54-dc8f1bc6c8d8" width="600"/
Downloads · 30 days
19
7% of all-time downloads
All-time downloads
272
Public
Repo size
25.2 MB
Likes
2
Public
Click a slice to open those files.
.pth25.2 MB · 100%
From the Hugging Face model README
llava-phi-3-mini-pretrain is a LLaVA projector pretrained from microsoft/Phi-3-mini-4k-instruct and CLIP-ViT-Large-patch14-336 on ShareGPT4V-PT dataset by XTuner.
The fine-tuned LLaVA model can be found on xtuner/llava-phi-3-mini.
@misc{2023xtuner,
title={XTuner: A Toolkit for Efficiently Fine-tuning LLM},
author={XTuner Contributors},
howpublished = {\url{https://github.com/InternLM/xtuner}},
year={2023}
}