Downloads · 30 days
9
31% of all-time downloads
weixinchen/GRATH-gradtruth
GRATH-gradtruth is a machine learning model from weixinchen. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft.
This is a gradually self-truthified model (with one iteration) proposed in the paper GRATH: Gradual Self-Truthifying for Large Language Models.
Downloads · 30 days
9
31% of all-time downloads
All-time downloads
29
Public
Repo size
25.7 MB
Likes
0
Public
Click a slice to open those files.
.bin25.2 MB · 91%
From the Hugging Face model README
This is a gradually self-truthified model (with one iteration) proposed in the paper GRATH: Gradual Self-Truthifying for Large Language Models.
Note: This model is applied with DPO twice. The reference model of DPO is set as the current base model.
The following bitsandbytes quantization config was used during training:
The following bitsandbytes quantization config was used during training:
PEFT 0.5.0
PEFT 0.5.0