Downloads · 30 days
0
RubyXZZZ/llama2-7b-self-aligned-final
llama2-7b-self-aligned-final is a machine learning model from RubyXZZZ. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Instruction-following model finetuned from meta-llama/Llama-2-7b-hf using QLoRA, implementing the self-alignment pipeline from ["Self-Alignment with Instruction Backtranslation"].
Downloads · 30 days
0
Access
Public
Updated Apr 2, 2026
Repo size
302 MB
Likes
0
Public
Click a slice to open those files.
.safetensors134 MB · 97%
From the Hugging Face model README
Instruction-following model finetuned from meta-llama/Llama-2-7b-hf using QLoRA, implementing the self-alignment pipeline from ["Self-Alignment with Instruction Backtranslation"].
This model is the final output of a 4-step self-alignment pipeline:
RubyXZZZ/lima-curated-backtranslation (self-curated high quality pairs, score >= 4)meta-llama/Llama-2-7b-hf@article{li2023self,
title={Self-alignment with instruction backtranslation},
journal={arXiv preprint arXiv:2308.06259},
year={2023}
}