Downloads · 30 days
3
16% of all-time downloads
GPRM/Llama-2-7b-LIMA-Alignment
Llama-2-7b-LIMA-Alignment is a machine learning model from GPRM. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
A straightforward implementation of the final Self-Alignment model, inspired by the paper "Self-Alignment with Instruction Backtranslation."
Downloads · 30 days
3
16% of all-time downloads
All-time downloads
19
Public
Parameters
6.7B
13.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.5 GB · 100%
From the Hugging Face model README
A straightforward implementation of the final Self-Alignment model, inspired by the paper "Self-Alignment with Instruction Backtranslation."
This model was fine-tuned on LLaMA-2-hf using self-curation LIMA dataset with high-quality sythetic (Instruction, Answer) pairs.
The fine-tuning process was conducted using LoRA, and the uploaded model is provided in its merged form.