Downloads · 30 days
0
hfl/chinese-mixtral-instruct-lora
chinese-mixtral-instruct-lora is a machine learning model from hfl. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
<p align="center" <a href="https://github.com/ymcui/Chinese-Mixtral"<img src="https://ymcui.com/images/chinese-mixtral-banner.png" width="600"/</a </p
Downloads · 30 days
0
Access
Public
Updated Mar 5, 2024
Repo size
2.5 GB
Likes
2
Public
Click a slice to open those files.
.safetensors2.5 GB · 100%
From the Hugging Face model README
Chinese Mixtral GitHub repository: https://github.com/ymcui/Chinese-Mixtral
This repository contains Chinese-Mixtral-Instruct-LoRA, which is further tuned with instruction data on Chinese-Mixtral, where Chinese-Mixtral is build on top of Mixtral-8x7B-v0.1.
Note: You must combine LoRA with the original Mixtral-8x7B-v0.1 to obtain full weight.
For full model, please see: https://huggingface.co/hfl/chinese-mixtral-instruct
For GGUF model (llama.cpp compatible), please see: https://huggingface.co/hfl/chinese-mixtral-instruct-gguf
If you have questions/issues regarding this model, please submit an issue through https://github.com/ymcui/Chinese-Mixtral/.
Please consider cite our paper if you use the resource of this repository. Paper link: https://arxiv.org/abs/2403.01851
@article{chinese-mixtral,
title={Rethinking LLM Language Adaptation: A Case Study on Chinese Mixtral},
author={Cui, Yiming and Yao, Xin},
journal={arXiv preprint arXiv:2403.01851},
url={https://arxiv.org/abs/2403.01851},
year={2024}
}