Downloads ยท 30 days
15
8% of all-time downloads
dunkiealex205/test-model
test-model is a text generation model from dunkiealex205. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.
This is the Full-Weight of WizardLM-13B V1.2 model, this model is trained from Llama-2 13b.
Downloads ยท 30 days
15
8% of all-time downloads
All-time downloads
190
Public
Repo size
52.1 GB
Likes
0
Public
Click a slice to open those files.
.bin26 GB ยท 100%
From the Hugging Face model README
This is the Full-Weight of WizardLM-13B V1.2 model, this model is trained from Llama-2 13b.
| Model | Checkpoint | Paper | HumanEval | MBPP | Demo | License |
|---|---|---|---|---|---|---|
| WizardCoder-Python-34B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-Python-34B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 73.2 | 61.2 | Demo | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama2</a> |
| WizardCoder-15B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-15B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 59.8 | 50.6 | -- | <a href="https://huggingface.co/spaces/bigcode/bigcode-model-license-agreement" target="_blank">OpenRAIL-M</a> |
| WizardCoder-Python-13B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-Python-13B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 64.0 | 55.6 | -- | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama2</a> |
| WizardCoder-Python-7B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-Python-7B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 55.5 | 51.6 | Demo | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama2</a> |
| WizardCoder-3B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-3B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 34.8 | 37.4 | -- | <a href="https://huggingface.co/spaces/bigcode/bigcode-model-license-agreement" target="_blank">OpenRAIL-M</a> |
| WizardCoder-1B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardCoder-1B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2306.08568" target="_blank">[WizardCoder]</a> | 23.8 | 28.6 | -- | <a href="https://huggingface.co/spaces/bigcode/bigcode-model-license-agreement" target="_blank">OpenRAIL-M</a> |
| Model | Checkpoint | Paper | GSM8k | MATH | Online Demo | License |
|---|---|---|---|---|---|---|
| WizardMath-70B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardMath-70B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2308.09583" target="_blank">[WizardMath]</a> | 81.6 | 22.7 | Demo | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama 2 </a> |
| WizardMath-13B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardMath-13B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2308.09583" target="_blank">[WizardMath]</a> | 63.9 | 14.0 | Demo | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama 2 </a> |
| WizardMath-7B-V1.0 | ๐ค <a href="https://huggingface.co/WizardLM/WizardMath-7B-V1.0" target="_blank">HF Link</a> | ๐ <a href="https://arxiv.org/abs/2308.09583" target="_blank">[WizardMath]</a> | 54.9 | 10.7 | Demo | <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama 2 </a> |
| <sup>Model</sup> | <sup>Checkpoint</sup> | <sup>Paper</sup> | <sup>MT-Bench</sup> | <sup>AlpacaEval</sup> | <sup>WizardEval</sup> | <sup>HumanEval</sup> | <sup>License</sup> |
|---|---|---|---|---|---|---|---|
| <sup>WizardLM-13B-V1.2</sup> | <sup>๐ค <a href="https://huggingface.co/WizardLM/WizardLM-13B-V1.2" target="_blank">HF Link</a> </sup> | <sup>7.06</sup> | <sup>89.17%</sup> | <sup>101.4% </sup> | <sup>36.6 pass@1</sup> | <sup> <a href="https://ai.meta.com/resources/models-and-libraries/llama-downloads/" target="_blank">Llama 2 License </a></sup> | |
| <sup>WizardLM-13B-V1.1</sup> | <sup> ๐ค <a href="https://huggingface.co/WizardLM/WizardLM-13B-V1.1" target="_blank">HF Link</a> </sup> | <sup>6.76</sup> | <sup>86.32%</sup> | <sup>99.3% </sup> | <sup>25.0 pass@1</sup> | <sup>Non-commercial</sup> | |
| <sup>WizardLM-30B-V1.0</sup> | <sup>๐ค <a href="https://huggingface.co/WizardLM/WizardLM-30B-V1.0" target="_blank">HF Link</a></sup> | <sup>7.01</sup> | <sup>97.8% </sup> | <sup>37.8 pass@1</sup> | <sup>Non-commercial</sup> | ||
| <sup>WizardLM-13B-V1.0</sup> | <sup>๐ค <a href="https://huggingface.co/WizardLM/WizardLM-13B-V1.0" target="_blank">HF Link</a> </sup> | <sup>6.35</sup> | <sup>75.31%</sup> | <sup>89.1% </sup> | <sup> 24.0 pass@1 </sup> | <sup>Non-commercial</sup> | |
| <sup>WizardLM-7B-V1.0 </sup> | <sup>๐ค <a href="https://huggingface.co/WizardLM/WizardLM-7B-V1.0" target="_blank">HF Link</a> </sup> | <sup> ๐ <a href="https://arxiv.org/abs/2304.12244" target="_blank">[WizardLM]</a> </sup> | <sup>78.0% </sup> | <sup>19.1 pass@1 </sup> | <sup> Non-commercial</sup> |
Repository: https://github.com/nlpxucan/WizardLM
Twitter:
โ<b>Note for model system prompts usage:</b>
<b>WizardLM</b> adopts the prompt format from <b>Vicuna</b> and supports multi-turn conversation. The prompt should be as following:
A chat between a curious user and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the user's questions. USER: Hi ASSISTANT: Hello.</s>USER: Who are you? ASSISTANT: I am WizardLM.</s>......
We provide the inference WizardLM demo code here.
Please cite the paper if you use the data or code from WizardLM.
@article{xu2023wizardlm,
title={Wizardlm: Empowering large language models to follow complex instructions},
author={Xu, Can and Sun, Qingfeng and Zheng, Kai and Geng, Xiubo and Zhao, Pu and Feng, Jiazhan and Tao, Chongyang and Jiang, Daxin},
journal={arXiv preprint arXiv:2304.12244},
year={2023}
}
โ<b>To commen concern about dataset:</b>
Recently, there have been clear changes in the open-source policy and regulations of our overall organization's code, data, and models.
Despite this, we have still worked hard to obtain opening the weights of the model first, but the data involves stricter auditing and is in review with our legal team .
Our researchers have no authority to publicly release them without authorization.
Thank you for your understanding.