Downloads · 30 days
26
6% of all-time downloads
DRXD1000/Phoenix-AWQ
Phoenix-AWQ is a text generation model from DRXD1000. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
<div align="center" <img src=https://cdn-uploads.huggingface.co/production/uploads/6474c16e7d131daf633db8ad/-mL8PSG00X2lEw1lb8E1Q.png </div
Downloads · 30 days
26
6% of all-time downloads
All-time downloads
441
Public
Parameters
7.2B
4.2 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors4.2 GB · 100%
How the weights are stored.
I327B · 96%
From the Hugging Face model README
| Bits | GS | AWQ Dataset | Seq Len |
|---|---|---|---|
| 4 | 128 | c4 | 4096 |
Phoenix is a model trained using Direct Preference Optimization (DPO) for the german language. Its training procedure follows the process of the alignment-handbook from Huggingface. In contrast to zephyr and notus this model has been trained using german instruction and dpo data. In detail, a german translation of HuggingFaceH4/ultrachat_200k and HuggingFaceH4/ultrafeedback_binarized were created in addition to a series of allready available instruction datasets. The LLM haoranxu/ALMA-13B was used for this. While the mistral model performs really well, it is not really suitable for the german language. Therefore we have used the fantastic LeoLM/leo-mistral-hessianai-7b. Thanks to the new type of training, Phoenix is not only able to compete with the Mistral model from LeoLM but also beats the Llama-70b-chat model in 2 mt-bench categories. This model wouldn't have been possible without the amazing work of Huggingface, LeoLM, openbnb, argilla, the Alma-Team and many others of the AI community. i would like to personally thank all AI researchers who make the training of such models possible
Phoenix beats the LeoLM-Mistral model in all categories except for coding and humanities. Additionally it also Beats LeoLM/Llama-2-70b-chat in roleplay and reasoning which shows the power of DPO.
{
"first_turn": 6.39375,
"second_turn": 5.1625,
"categories": {
"writing": 7.45,
"roleplay": 7.9,
"reasoning": 4.3,
"math": 3.25,
"coding": 2.5,
"extraction": 5.9,
"stem": 7.125,
"humanities": 7.8
},
"average": 5.778124999999999
}
Florian Leurer compared Phoenix to other LLMs. Check it out here:
LeoLM/leo-mistral-hessianai-7bPHOENIX: Open-Source Language Adaption for Direct Preference OptimizationWe used a VM with 8 x A100 80GB hosted in Runpods.io.
We used a new translated version of HuggingFaceH4/ultrachat_200k, and argilla/ultrafeedback-binarized-preferences.
The data used for training will be made public after additional quality inspection.
We use the same prompt template as HuggingFaceH4/zephyr-7b-beta:
<|system|>
</s>
<|user|>
{prompt}</s>
<|assistant|>
It is also possible to use the model in a multi-turn setup
<|system|>
</s>
<|user|>
{prompt_1}</s>
<|assistant|>
{answer_1}</s>
<|user|>
{prompt_2}</s>
<|assistant|>
You will first need to install transformers and accelerate (just to ease the device placement), then you can run any of the following:
generateimport torch
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("DRXD1000/Phoenix-AWQ", torch_dtype=torch.bfloat16, device_map="auto")
tokenizer = AutoTokenizer.from_pretrained("DRXD1000/Phoenix-AWQ")
prompt = [
{
"role": "system",
"content": "", #Not recommended. Phoenix does not react well on system prompts
},
{"role": "user", "content": "Erkläre mir was KI ist"},
]
inputs = tokenizer.apply_chat_template(prompt, return_tensors="pt").to("cuda")
outputs = model.generate(inputs, num_return_sequences=1, max_new_tokens=256, do_sample=True, temperature=0.7, top_k=50, top_p=0.95)
response = tokenizer.decode(outputs[0], skip_special_tokens=True)
As with all LLMs, the potential outputs of DRXD1000/Phoenix cannot be predicted
in advance, and the model may in some instances produce inaccurate, biased or other objectionable responses
to user prompts. Therefore, before deploying any applications of DRXD1000/Phoenix, developers should
perform safety testing and tuning tailored to their specific applications of the model.
Please see Meta's Responsible Use Guide.
The following hyperparameters were used during training:
@misc{uhlig2024phoenix,
title={PHOENIX: Open-Source Language Adaption for Direct Preference Optimization},
author={Matthias Uhlig and Sigurd Schacht and Sudarshan Kamath Barkur},
year={2024},
eprint={2401.10580},
archivePrefix={arXiv},
primaryClass={cs.CL}
}