Downloads · 30 days
426
10% of all-time downloads
QuantFactory/Azure_Dusk-v0.2-GGUF
Azure_Dusk-v0.2-GGUF is a text generation model from QuantFactory. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
Downloads · 30 days
426
10% of all-time downloads
All-time downloads
4.4K
Public
Repo size
111 GB
Likes
2
Public
Click a slice to open those files.
.gguf111 GB · 100%
From the Hugging Face model README
This is quantized version of Epiculous/Azure_Dusk-v0.2 created using llama.cpp

Following up on Crimson_Dawn-v0.2 we have Azure_Dusk-v0.2! Training on Mistral-Nemo-Base-2407 this time I've added significantly more data, as well as trained using RSLoRA as opposed to regular LoRA. Another key change is training on ChatML as opposed to Mistral Formatting.
<strong>full</strong> / exl2 / gguf
The v0.2 models are trained on ChatML, the prompting structure goes a little something like this:
<|im_start|>user
Hi there!<|im_end|>
<|im_start|>assistant
Nice to meet you!<|im_end|>
<|im_start|>user
Can I ask a question?<|im_end|>
<|im_start|>assistant
The v0.2 models are trained on ChatML, please use that Context and Instruct template.
Spicy_Temp <br/> Violet_Twilight-Nitral-Special <br/>
Training was done twice over 2 epochs each on two 2x NVIDIA A6000 GPUs using LoRA. A two-phased approach was used in which the base model was trained 2 epochs on RP data, the LoRA was then applied to base. Finally, the new modified base was trained 2 epochs on instruct, and the new instruct LoRA was applied to the modified base, resulting in what you see here.