Downloads · 30 days
17
33% of all-time downloads
colable/llama2-ko-DPO
llama2-ko-DPO is a text generation model from colable. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
This is an Korean Model based on [beomi/open-llama-2-ko-7b]
Downloads · 30 days
17
33% of all-time downloads
All-time downloads
52
Public
Parameters
6.9B
13.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.7 GB · 100%
From the Hugging Face model README
This is an Korean Model based on
Dataset is modified from
Parameters
learning_rate: float = 3e-4
lr_scheduler: str = "cosine"
warmup_ratio: float = 0.1
lora_r: int = 16
lora_alpha: int = 16
lora_dropout: float = 0.05
optim='paged_adamw_32bit'
bf16=True